LLM Perks

Free long-context models

Models with a context window of 256K tokens or more — for RAG without chunking, whole-repository reviews, and long documents.

65 models11 providersup to 10.5M contextbest: GLM 5.3Updated:
Read more

Free routes often cap context below the model's maximum. The column shows the provider-specific limit.

Size
65 found
# Model ↕ AA Index ↕ GPQA ↕ Context ↕ Size ↕ Free at ↕
1
Llama 4 Scout
Meta · 109B · 17B
VisionTool useOpen weights
8.1 59% 10.5M Frontier Together AI
2
Gemini 1.5 Pro
Google · Legacy
Vision
7.9 59% 2.1M Frontier Google AI Studio
3
LongCat 2.0
Meituan
ReasoningTool useOpen weights
19.1 78% 1M Frontier Nous Portal
4
GLM 5.3
Z.ai (Zhipu AI)
ReasoningCodingTool useOpen weights
44.8 92% 1M Frontier Together AI, NVIDIA Build
5
GLM 5.2
Z.ai (Zhipu AI)
ReasoningCodingTool useOpen weights
33.7 90% 1M Frontier Together AI
6
Inkling Small
Thinking Machines Lab
ReasoningOpen weights
27.8 90% 1M Small OpenRouter
7
Nemotron 3 Ultra
NVIDIA · 550B · 55B
ReasoningCodingTool useOpen weights
22.9 87% 1M Frontier OpenRouter, NVIDIA Build, UnoRouter
8
Inkling
Thinking Machines Lab
ReasoningOpen weights
25.0 87% 1M Small OpenRouter
9
Llama 4 Maverick Instruct Fp4
Meta · 400B · 17B
VisionTool useOpen weights
— — 1M Frontier Together AI
10
MiniMax M1 80K
MiniMax
ReasoningTool useOpen weights
11.7 70% 1M Frontier Together AI
11
MiniMax M1 40K
MiniMax
ReasoningTool useOpen weights
10.0 68% 1M Frontier Together AI
12
Laguna S 2.1
Poolside
ReasoningCoding
— — 1M Small OpenRouter, Vercel AI Gateway, Nous Portal, UnoRouter
13
Gemini 2.0 Flash
Google
VisionTool use
8.9 62% 1M Mid-size Google AI Studio
14
Gemini 2.0 Flash Lite
Google
Vision
7.3 54% 1M Small Google AI Studio
15
Gemini 1.5 Flash
Google · Legacy
Vision
7.1 46% 1M Mid-size Google AI Studio
16
Mimo V2.6 Pro
New
Open weights
46.3 — 1M Frontier UnoRouter
17
Kimi K3
Moonshot AI
VisionReasoningCodingTool useOpen weights
43.6 94% 1M Frontier NVIDIA Build, UnoRouter
18
GLM 5.3 Flash
Z.ai (Zhipu AI)
ReasoningCodingTool useOpen weights
41.8 91% 1M Mid-size OrcaRouter, NVIDIA Build, UnoRouter
19
Qwen 3.8 Flash Next
Alibaba Cloud · New
ReasoningCodingTool useOpen weights
39.8 92% 1M Mid-size UnoRouter
20
DeepSeek V4 Flash
DeepSeek
ReasoningCodingTool useOpen weights
34.3 91% 1M Frontier OrcaRouter, UnoRouter
21
Gemini 3.6 Flash
Google
VisionReasoningCodingTool use
34.0 93% 1M Mid-size UnoRouter
22
Gemini 3.5 Flash Lite
Google
VisionReasoningCoding
22.2 84% 1M Small UnoRouter
23
Nemotron 3.5 Lightning
NVIDIA · 30B · 3B
ReasoningTool useOpen weights
12.9 74% 1M Mid-size OpenRouter, NVIDIA Build, UnoRouter
24
GLM 5.3 Flash Search
Z.ai (Zhipu AI)
ReasoningCodingTool useOpen weights
— — 1M Mid-size UnoRouter
25
Space Bunny Alpha
Stealth model · New
— — 1M Small OpenRouter, UnoRouter
26 — — 1M Small UnoRouter
27
Solar Pro 4
Upstage
ReasoningTool use
28.2 89% 524K Frontier Nous Portal
28
— — 524K Small UnoRouter
29 — — 512K Frontier UnoRouter
30
Dots3-Note Preview
Dots Studio
— — 512K Small OpenRouter, UnoRouter
31 — — 512K Small UnoRouter
32 — — 512K Small UnoRouter
33
Qwen3.8 27B
Alibaba Cloud · 27B
ReasoningCodingTool useOpen weights
33.7 91% 262K Mid-size OpenRouter, Groq, UnoRouter
34
Ling 3.0 Flash Fin
inclusionAI (Ant Group)
Open weights
22.6 — 262K Small OpenRouter, Nous Portal, UnoRouter
35
Gemma 4 31B
Google DeepMind · 31B
VisionTool useOpen weights
19.0 86% 262K Mid-size OpenRouter, Together AI, NVIDIA Build, UnoRouter
36
Gemma 4 26B A4B
Google DeepMind · 26B · 4B
VisionTool useOpen weights
16.7 79% 262K Mid-size OpenRouter, Together AI, UnoRouter
37
Step 3.7 Flash
StepFun
ReasoningOpen weights
19.5 81% 262K Small Nous Portal
38
Qwen 3.5 35B
Alibaba Cloud · 35B · 3B
ReasoningCodingTool useOpen weights
19.3 85% 262K Mid-size Together AI
39
Qwen 3.6 35B
Alibaba Cloud · 35B · 3B
ReasoningCodingTool useOpen weights
18.2 84% 262K Mid-size Together AI
40
Qwen 3.5 122B
Alibaba Cloud · 122B · 10B
ReasoningCodingTool useOpen weights
17.7 83% 262K Frontier Together AI
41
Nemotron 3 Super
NVIDIA · 120B · 12B
ReasoningCodingTool useOpen weights
12.8 80% 262K Frontier OpenRouter, NVIDIA Build, UnoRouter
42
Nemotron 3 Nano Omni
NVIDIA · 30B · 3B
VisionReasoningTool useOpen weights
10.3 47% 262K Mid-size +1OpenRouter, Together AI, NVIDIA Build, TokenRouter, UnoRouter
43
Gemma 4 12B
Google DeepMind · 12B
VisionTool useOpen weights
14.2 75% 262K Mid-size Together AI
44
Qwen 3.5 9B
Alibaba Cloud · 9B
ReasoningCodingOpen weights
13.7 81% 262K Small Together AI
45
Qwen3 VL 235B
Alibaba Cloud · 235B · 22B
VisionReasoningTool useOpen weights
13.4 77% 262K Frontier Together AI
46
Nvidia Nemotron 3 Super
NVIDIA · 120B · 12B
ReasoningCodingTool useOpen weights
12.8 80% 262K Frontier Together AI
47
Qwen3 235B
Alibaba Cloud · 235B · 22B
ReasoningCodingTool useOpen weights
12.7 79% 262K Frontier Together AI
48
North Mini Code
Cohere
CodingOpen weights
9.9 76% 262K Small OpenRouter, UnoRouter
49
Qwen3 30B
Alibaba Cloud · 30B · 3B
ReasoningTool useOpen weights
9.8 71% 262K Mid-size Together AI
50
Qwen3 Coder 30B
Alibaba Cloud · 30B · 3B
ReasoningCodingTool useOpen weights
9.6 52% 262K Frontier Together AI
51
Ling 3.0 Flash Sante
inclusionAI (Ant Group)
Open weights
— — 262K Small OpenRouter, Vercel AI Gateway, Nous Portal, UnoRouter
52
Laguna XS 2.1
Poolside
ReasoningCoding
— — 262K Small OpenRouter, NVIDIA Build, Nous Portal, UnoRouter
53
DiffusionGemma 26B
Google DeepMind · 26B · 4B
Tool useOpen weights
9.5 67% 262K Mid-size UnoRouter
54 — — 262K Mid-size Together AI
55
Holo3 35B A3B
35B · 3B
— — 262K Mid-size Together AI
56 — — 262K Mid-size UnoRouter
57
Qwen SEA-LION V4.5 27B
Alibaba Cloud · 27B
Tool useOpen weights
— — 262K Mid-size UnoRouter
58
Nvidia Nemotron 3 Nano
NVIDIA · 30B · 3B
ReasoningTool useOpen weights
8.9 76% 262K Mid-size Together AI
59
Qwen3 4B
Alibaba Cloud · 4B
ReasoningOpen weights
8.8 67% 262K Small Together AI
60
Qwen 3.5 2B
Alibaba Cloud · 2B
ReasoningCodingOpen weights
6.9 46% 262K Small Together AI
61 — — 262K Small UnoRouter
62
Mistral Large 3 675B
Mistral AI · 675B
CodingTool useOpen weights
— — 256K Frontier UnoRouter
63
Mistral Small
Mistral AI · 24B
Tool useOpen weights
5.8 38% 256K Mid-size Mistral AI, UnoRouter
64
Codestral Mamba
Mistral AI
Coding
— — 256K Small Mistral AI
65
Codestral
Mistral AI
Coding
— — 256K Small UnoRouter

Benchmarks: Artificial Analysis (September 24, 2026). A dash means the model is not on their leaderboard; we never estimate scores ourselves.

FAQ

What is the best free model for long context?

GLM 5.3 currently leads — free at Together AI, NVIDIA Build. Next come Mimo V2.6 Pro, Kimi K3, GLM 5.3 Flash.

See also