LLM Perks

Mistral Nemo Minitron 8B 8K

Mistral AI · All Mistral models
Free right now Small
Open weights

Mistral Nemo Minitron 8B 8K is a language model with 8B parameters from Mistral AI. It is currently free at 1 provider with up to 8K tokens of context. Its weights are open, so it can also run locally.

Developer
Mistral AI
Parameters
8B
Context
8K tokens
Type
Text
License
Open weights

Where to use Mistral Nemo Minitron 8B 8K for free

Each row is a separate free endpoint with its own limits.

ProviderContextLimitCard
NVIDIA Build
nvidia/mistral-nemo-minitron-8b-8k-instruct
8K Free serverless inference for development; limits and model availability may change. No Open

Quick start

curl https://integrate.api.nvidia.com/v1/chat/completions \
  -H "Authorization: Bearer $NVIDIA_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "nvidia/mistral-nemo-minitron-8b-8k-instruct",
    "messages": [{"role": "user", "content": "Hello!"}]
  }'

Example for NVIDIA Build. Any OpenAI-compatible client works: Cursor, Cline, OpenCode, Open WebUI.

Benchmarks

This model is not on the Artificial Analysis leaderboard yet. We do not publish our own estimates.

Similar free models

FAQ

Is Mistral Nemo Minitron 8B 8K free?

Yes. As of September 26, 2026, Mistral Nemo Minitron 8B 8K is free at NVIDIA Build. Each provider has its own limits — see the table above.

What is the context window of Mistral Nemo Minitron 8B 8K?

The largest context among free routes is 8,192 tokens. Some providers cap it lower.

How do I call Mistral Nemo Minitron 8B 8K via API?

Use any OpenAI-compatible client with base URL https://integrate.api.nvidia.com/v1, model ID nvidia/mistral-nemo-minitron-8b-8k-instruct, and a NVIDIA Build key.

Is Mistral Nemo Minitron 8B 8K open-weight?

Yes, the weights are published — you can download and run it locally, for example with Ollama or LM Studio.