LLM Perks

Modèles à long contexte gratuits

Modèles avec une fenêtre de contexte de 256K tokens ou plus — pour du RAG sans découpage, la revue de dépôts entiers et les longs documents.

65 modèles11 fournisseursjusqu'à 10.5M contextemeilleur : GLM 5.3Mis à jour:
En savoir plus

Les routes gratuites plafonnent souvent le contexte en dessous du maximum du modèle. La colonne indique la limite propre au fournisseur.

Taille
65 trouvé(s)
# Modèle ↕ Indice AA ↕ GPQA ↕ Contexte ↕ Taille ↕ Gratuit chez ↕
1
Llama 4 Scout
Meta · 109B · 17B
VisionAppel d'outilsPoids ouverts
8.1 59% 10.5M Très grand Together AI
2
Gemini 1.5 Pro
Google · Ancienne génération
Vision
7.9 59% 2.1M Très grand Google AI Studio
3
LongCat 2.0
Meituan
RaisonnementAppel d'outilsPoids ouverts
19.1 78% 1M Très grand Nous Portal
4
GLM 5.3
Z.ai (Zhipu AI)
RaisonnementCodeAppel d'outilsPoids ouverts
44.8 92% 1M Très grand Together AI, NVIDIA Build
5
GLM 5.2
Z.ai (Zhipu AI)
RaisonnementCodeAppel d'outilsPoids ouverts
33.7 90% 1M Très grand Together AI
6
Inkling Small
Thinking Machines Lab
RaisonnementPoids ouverts
27.8 90% 1M Petit OpenRouter
7
Nemotron 3 Ultra
NVIDIA · 550B · 55B
RaisonnementCodeAppel d'outilsPoids ouverts
22.9 87% 1M Très grand OpenRouter, NVIDIA Build, UnoRouter
8
Inkling
Thinking Machines Lab
RaisonnementPoids ouverts
25.0 87% 1M Petit OpenRouter
9
Llama 4 Maverick Instruct Fp4
Meta · 400B · 17B
VisionAppel d'outilsPoids ouverts
— — 1M Très grand Together AI
10
MiniMax M1 80K
MiniMax
RaisonnementAppel d'outilsPoids ouverts
11.7 70% 1M Très grand Together AI
11
MiniMax M1 40K
MiniMax
RaisonnementAppel d'outilsPoids ouverts
10.0 68% 1M Très grand Together AI
12
Laguna S 2.1
Poolside
RaisonnementCode
— — 1M Petit OpenRouter, Vercel AI Gateway, Nous Portal, UnoRouter
13
Gemini 2.0 Flash
Google
VisionAppel d'outils
8.9 62% 1M Taille moyenne Google AI Studio
14
Gemini 2.0 Flash Lite
Google
Vision
7.3 54% 1M Petit Google AI Studio
15
Gemini 1.5 Flash
Google · Ancienne génération
Vision
7.1 46% 1M Taille moyenne Google AI Studio
16
Mimo V2.6 Pro
New
Poids ouverts
46.3 — 1M Très grand UnoRouter
17
Kimi K3
Moonshot AI
VisionRaisonnementCodeAppel d'outilsPoids ouverts
43.6 94% 1M Très grand NVIDIA Build, UnoRouter
18
GLM 5.3 Flash
Z.ai (Zhipu AI)
RaisonnementCodeAppel d'outilsPoids ouverts
41.8 91% 1M Taille moyenne OrcaRouter, NVIDIA Build, UnoRouter
19
Qwen 3.8 Flash Next
Alibaba Cloud · New
RaisonnementCodeAppel d'outilsPoids ouverts
39.8 92% 1M Taille moyenne UnoRouter
20
DeepSeek V4 Flash
DeepSeek
RaisonnementCodeAppel d'outilsPoids ouverts
34.3 91% 1M Très grand OrcaRouter, UnoRouter
21
Gemini 3.6 Flash
Google
VisionRaisonnementCodeAppel d'outils
34.0 93% 1M Taille moyenne UnoRouter
22
Gemini 3.5 Flash Lite
Google
VisionRaisonnementCode
22.2 84% 1M Petit UnoRouter
23
Nemotron 3.5 Lightning
NVIDIA · 30B · 3B
RaisonnementAppel d'outilsPoids ouverts
12.9 74% 1M Taille moyenne OpenRouter, NVIDIA Build, UnoRouter
24
GLM 5.3 Flash Search
Z.ai (Zhipu AI)
RaisonnementCodeAppel d'outilsPoids ouverts
— — 1M Taille moyenne UnoRouter
25
Space Bunny Alpha
Stealth model · New
— — 1M Petit OpenRouter, UnoRouter
26 — — 1M Petit UnoRouter
27
Solar Pro 4
Upstage
RaisonnementAppel d'outils
28.2 89% 524K Très grand Nous Portal
28
— — 524K Petit UnoRouter
29 — — 512K Très grand UnoRouter
30
Dots3-Note Preview
Dots Studio
— — 512K Petit OpenRouter, UnoRouter
31 — — 512K Petit UnoRouter
32 — — 512K Petit UnoRouter
33
Qwen3.8 27B
Alibaba Cloud · 27B
RaisonnementCodeAppel d'outilsPoids ouverts
33.7 91% 262K Taille moyenne OpenRouter, Groq, UnoRouter
34
Ling 3.0 Flash Fin
inclusionAI (Ant Group)
Poids ouverts
22.6 — 262K Petit OpenRouter, Nous Portal, UnoRouter
35
Gemma 4 31B
Google DeepMind · 31B
VisionAppel d'outilsPoids ouverts
19.0 86% 262K Taille moyenne OpenRouter, Together AI, NVIDIA Build, UnoRouter
36
Gemma 4 26B A4B
Google DeepMind · 26B · 4B
VisionAppel d'outilsPoids ouverts
16.7 79% 262K Taille moyenne OpenRouter, Together AI, UnoRouter
37
Step 3.7 Flash
StepFun
RaisonnementPoids ouverts
19.5 81% 262K Petit Nous Portal
38
Qwen 3.5 35B
Alibaba Cloud · 35B · 3B
RaisonnementCodeAppel d'outilsPoids ouverts
19.3 85% 262K Taille moyenne Together AI
39
Qwen 3.6 35B
Alibaba Cloud · 35B · 3B
RaisonnementCodeAppel d'outilsPoids ouverts
18.2 84% 262K Taille moyenne Together AI
40
Qwen 3.5 122B
Alibaba Cloud · 122B · 10B
RaisonnementCodeAppel d'outilsPoids ouverts
17.7 83% 262K Très grand Together AI
41
Nemotron 3 Super
NVIDIA · 120B · 12B
RaisonnementCodeAppel d'outilsPoids ouverts
12.8 80% 262K Très grand OpenRouter, NVIDIA Build, UnoRouter
42
Nemotron 3 Nano Omni
NVIDIA · 30B · 3B
VisionRaisonnementAppel d'outilsPoids ouverts
10.3 47% 262K Taille moyenne +1OpenRouter, Together AI, NVIDIA Build, TokenRouter, UnoRouter
43
Gemma 4 12B
Google DeepMind · 12B
VisionAppel d'outilsPoids ouverts
14.2 75% 262K Taille moyenne Together AI
44
Qwen 3.5 9B
Alibaba Cloud · 9B
RaisonnementCodePoids ouverts
13.7 81% 262K Petit Together AI
45
Qwen3 VL 235B
Alibaba Cloud · 235B · 22B
VisionRaisonnementAppel d'outilsPoids ouverts
13.4 77% 262K Très grand Together AI
46
Nvidia Nemotron 3 Super
NVIDIA · 120B · 12B
RaisonnementCodeAppel d'outilsPoids ouverts
12.8 80% 262K Très grand Together AI
47
Qwen3 235B
Alibaba Cloud · 235B · 22B
RaisonnementCodeAppel d'outilsPoids ouverts
12.7 79% 262K Très grand Together AI
48
North Mini Code
Cohere
CodePoids ouverts
9.9 76% 262K Petit OpenRouter, UnoRouter
49
Qwen3 30B
Alibaba Cloud · 30B · 3B
RaisonnementAppel d'outilsPoids ouverts
9.8 71% 262K Taille moyenne Together AI
50
Qwen3 Coder 30B
Alibaba Cloud · 30B · 3B
RaisonnementCodeAppel d'outilsPoids ouverts
9.6 52% 262K Très grand Together AI
51
Ling 3.0 Flash Sante
inclusionAI (Ant Group)
Poids ouverts
— — 262K Petit OpenRouter, Vercel AI Gateway, Nous Portal, UnoRouter
52
Laguna XS 2.1
Poolside
RaisonnementCode
— — 262K Petit OpenRouter, NVIDIA Build, Nous Portal, UnoRouter
53
DiffusionGemma 26B
Google DeepMind · 26B · 4B
Appel d'outilsPoids ouverts
9.5 67% 262K Taille moyenne UnoRouter
54 — — 262K Taille moyenne Together AI
55
Holo3 35B A3B
35B · 3B
— — 262K Taille moyenne Together AI
56 — — 262K Taille moyenne UnoRouter
57
Qwen SEA-LION V4.5 27B
Alibaba Cloud · 27B
Appel d'outilsPoids ouverts
— — 262K Taille moyenne UnoRouter
58
Nvidia Nemotron 3 Nano
NVIDIA · 30B · 3B
RaisonnementAppel d'outilsPoids ouverts
8.9 76% 262K Taille moyenne Together AI
59
Qwen3 4B
Alibaba Cloud · 4B
RaisonnementPoids ouverts
8.8 67% 262K Petit Together AI
60
Qwen 3.5 2B
Alibaba Cloud · 2B
RaisonnementCodePoids ouverts
6.9 46% 262K Petit Together AI
61 — — 262K Petit UnoRouter
62
Mistral Large 3 675B
Mistral AI · 675B
CodeAppel d'outilsPoids ouverts
— — 256K Très grand UnoRouter
63
Mistral Small
Mistral AI · 24B
Appel d'outilsPoids ouverts
5.8 38% 256K Taille moyenne Mistral AI, UnoRouter
64
Codestral Mamba
Mistral AI
Code
— — 256K Petit Mistral AI
65
Codestral
Mistral AI
Code
— — 256K Petit UnoRouter

Benchmarks : Artificial Analysis (24 septembre 2026). Un tiret signifie que le modèle ne figure pas dans leur classement ; nous n'estimons jamais les scores nous-mêmes.

FAQ

Quel est le meilleur modèle gratuit pour long contexte ?

GLM 5.3 est actuellement en tête — gratuit chez Together AI, NVIDIA Build. Viennent ensuite Mimo V2.6 Pro, Kimi K3, GLM 5.3 Flash.

Voir aussi