LLM Perks

Nemotron 4 340B

Free right now Frontier
Open weights

Nemotron 4 340B is a language model with 340B parameters from NVIDIA. It is currently free at 1 provider with up to — tokens of context. Its weights are open, so it can also run locally.

Developer
NVIDIA
Parameters
340B
Type
Text
License
Open weights

Where to use Nemotron 4 340B for free

Each row is a separate free endpoint with its own limits.

ProviderContextLimitCard
NVIDIA Build
nvidia/nemotron-4-340b-instruct
— Free serverless inference for development; limits and model availability may change. No Open

Quick start

curl https://integrate.api.nvidia.com/v1/chat/completions \
  -H "Authorization: Bearer $NVIDIA_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "nvidia/nemotron-4-340b-instruct",
    "messages": [{"role": "user", "content": "Hello!"}]
  }'

Example for NVIDIA Build. Any OpenAI-compatible client works: Cursor, Cline, OpenCode, Open WebUI.

Benchmarks

This model is not on the Artificial Analysis leaderboard yet. We do not publish our own estimates.

Similar free models

FAQ

Is Nemotron 4 340B free?

Yes. As of September 26, 2026, Nemotron 4 340B is free at NVIDIA Build. Each provider has its own limits — see the table above.

How do I call Nemotron 4 340B via API?

Use any OpenAI-compatible client with base URL https://integrate.api.nvidia.com/v1, model ID nvidia/nemotron-4-340b-instruct, and a NVIDIA Build key.

Is Nemotron 4 340B open-weight?

Yes, the weights are published — you can download and run it locally, for example with Ollama or LM Studio.