LLM Perks

Kostenlose Modelle mit langem Kontext

Modelle mit einem Kontextfenster von 256K Token oder mehr — für RAG ohne Chunking, Reviews ganzer Repositories und lange Dokumente.

65 Modelle11 Anbieterbis zu 10.5M Kontextbeste: GLM 5.3Aktualisiert:
Weiterlesen

Kostenlose Routen begrenzen den Kontext oft unter das Maximum des Modells. Die Spalte zeigt das anbieterspezifische Limit.

Größe
65 gefunden
# Modell ↕ AA Index ↕ GPQA ↕ Kontext ↕ Größe ↕ Kostenlos bei ↕
1
Llama 4 Scout
Meta · 109B · 17B
VisionTool-NutzungOffene Gewichte
8.1 59% 10.5M Sehr groß Together AI
2
Gemini 1.5 Pro
Google · Legacy
Vision
7.9 59% 2.1M Sehr groß Google AI Studio
3
LongCat 2.0
Meituan
ReasoningTool-NutzungOffene Gewichte
19.1 78% 1M Sehr groß Nous Portal
4
GLM 5.3
Z.ai (Zhipu AI)
ReasoningCodingTool-NutzungOffene Gewichte
44.8 92% 1M Sehr groß Together AI, NVIDIA Build
5
GLM 5.2
Z.ai (Zhipu AI)
ReasoningCodingTool-NutzungOffene Gewichte
33.7 90% 1M Sehr groß Together AI
6
Inkling Small
Thinking Machines Lab
ReasoningOffene Gewichte
27.8 90% 1M Klein OpenRouter
7
Nemotron 3 Ultra
NVIDIA · 550B · 55B
ReasoningCodingTool-NutzungOffene Gewichte
22.9 87% 1M Sehr groß OpenRouter, NVIDIA Build, UnoRouter
8
Inkling
Thinking Machines Lab
ReasoningOffene Gewichte
25.0 87% 1M Klein OpenRouter
9
Llama 4 Maverick Instruct Fp4
Meta · 400B · 17B
VisionTool-NutzungOffene Gewichte
— — 1M Sehr groß Together AI
10
MiniMax M1 80K
MiniMax
ReasoningTool-NutzungOffene Gewichte
11.7 70% 1M Sehr groß Together AI
11
MiniMax M1 40K
MiniMax
ReasoningTool-NutzungOffene Gewichte
10.0 68% 1M Sehr groß Together AI
12
Laguna S 2.1
Poolside
ReasoningCoding
— — 1M Klein OpenRouter, Vercel AI Gateway, Nous Portal, UnoRouter
13
Gemini 2.0 Flash
Google
VisionTool-Nutzung
8.9 62% 1M Mittelgroß Google AI Studio
14
Gemini 2.0 Flash Lite
Google
Vision
7.3 54% 1M Klein Google AI Studio
15
Gemini 1.5 Flash
Google · Legacy
Vision
7.1 46% 1M Mittelgroß Google AI Studio
16
Mimo V2.6 Pro
New
Offene Gewichte
46.3 — 1M Sehr groß UnoRouter
17
Kimi K3
Moonshot AI
VisionReasoningCodingTool-NutzungOffene Gewichte
43.6 94% 1M Sehr groß NVIDIA Build, UnoRouter
18
GLM 5.3 Flash
Z.ai (Zhipu AI)
ReasoningCodingTool-NutzungOffene Gewichte
41.8 91% 1M Mittelgroß OrcaRouter, NVIDIA Build, UnoRouter
19
Qwen 3.8 Flash Next
Alibaba Cloud · New
ReasoningCodingTool-NutzungOffene Gewichte
39.8 92% 1M Mittelgroß UnoRouter
20
DeepSeek V4 Flash
DeepSeek
ReasoningCodingTool-NutzungOffene Gewichte
34.3 91% 1M Sehr groß OrcaRouter, UnoRouter
21
Gemini 3.6 Flash
Google
VisionReasoningCodingTool-Nutzung
34.0 93% 1M Mittelgroß UnoRouter
22
Gemini 3.5 Flash Lite
Google
VisionReasoningCoding
22.2 84% 1M Klein UnoRouter
23
Nemotron 3.5 Lightning
NVIDIA · 30B · 3B
ReasoningTool-NutzungOffene Gewichte
12.9 74% 1M Mittelgroß OpenRouter, NVIDIA Build, UnoRouter
24
GLM 5.3 Flash Search
Z.ai (Zhipu AI)
ReasoningCodingTool-NutzungOffene Gewichte
— — 1M Mittelgroß UnoRouter
25
Space Bunny Alpha
Stealth model · New
— — 1M Klein OpenRouter, UnoRouter
26 — — 1M Klein UnoRouter
27
Solar Pro 4
Upstage
ReasoningTool-Nutzung
28.2 89% 524K Sehr groß Nous Portal
28
— — 524K Klein UnoRouter
29 — — 512K Sehr groß UnoRouter
30
Dots3-Note Preview
Dots Studio
— — 512K Klein OpenRouter, UnoRouter
31 — — 512K Klein UnoRouter
32 — — 512K Klein UnoRouter
33
Qwen3.8 27B
Alibaba Cloud · 27B
ReasoningCodingTool-NutzungOffene Gewichte
33.7 91% 262K Mittelgroß OpenRouter, Groq, UnoRouter
34
Ling 3.0 Flash Fin
inclusionAI (Ant Group)
Offene Gewichte
22.6 — 262K Klein OpenRouter, Nous Portal, UnoRouter
35
Gemma 4 31B
Google DeepMind · 31B
VisionTool-NutzungOffene Gewichte
19.0 86% 262K Mittelgroß OpenRouter, Together AI, NVIDIA Build, UnoRouter
36
Gemma 4 26B A4B
Google DeepMind · 26B · 4B
VisionTool-NutzungOffene Gewichte
16.7 79% 262K Mittelgroß OpenRouter, Together AI, UnoRouter
37
Step 3.7 Flash
StepFun
ReasoningOffene Gewichte
19.5 81% 262K Klein Nous Portal
38
Qwen 3.5 35B
Alibaba Cloud · 35B · 3B
ReasoningCodingTool-NutzungOffene Gewichte
19.3 85% 262K Mittelgroß Together AI
39
Qwen 3.6 35B
Alibaba Cloud · 35B · 3B
ReasoningCodingTool-NutzungOffene Gewichte
18.2 84% 262K Mittelgroß Together AI
40
Qwen 3.5 122B
Alibaba Cloud · 122B · 10B
ReasoningCodingTool-NutzungOffene Gewichte
17.7 83% 262K Sehr groß Together AI
41
Nemotron 3 Super
NVIDIA · 120B · 12B
ReasoningCodingTool-NutzungOffene Gewichte
12.8 80% 262K Sehr groß OpenRouter, NVIDIA Build, UnoRouter
42
Nemotron 3 Nano Omni
NVIDIA · 30B · 3B
VisionReasoningTool-NutzungOffene Gewichte
10.3 47% 262K Mittelgroß +1OpenRouter, Together AI, NVIDIA Build, TokenRouter, UnoRouter
43
Gemma 4 12B
Google DeepMind · 12B
VisionTool-NutzungOffene Gewichte
14.2 75% 262K Mittelgroß Together AI
44
Qwen 3.5 9B
Alibaba Cloud · 9B
ReasoningCodingOffene Gewichte
13.7 81% 262K Klein Together AI
45
Qwen3 VL 235B
Alibaba Cloud · 235B · 22B
VisionReasoningTool-NutzungOffene Gewichte
13.4 77% 262K Sehr groß Together AI
46
Nvidia Nemotron 3 Super
NVIDIA · 120B · 12B
ReasoningCodingTool-NutzungOffene Gewichte
12.8 80% 262K Sehr groß Together AI
47
Qwen3 235B
Alibaba Cloud · 235B · 22B
ReasoningCodingTool-NutzungOffene Gewichte
12.7 79% 262K Sehr groß Together AI
48
North Mini Code
Cohere
CodingOffene Gewichte
9.9 76% 262K Klein OpenRouter, UnoRouter
49
Qwen3 30B
Alibaba Cloud · 30B · 3B
ReasoningTool-NutzungOffene Gewichte
9.8 71% 262K Mittelgroß Together AI
50
Qwen3 Coder 30B
Alibaba Cloud · 30B · 3B
ReasoningCodingTool-NutzungOffene Gewichte
9.6 52% 262K Sehr groß Together AI
51
Ling 3.0 Flash Sante
inclusionAI (Ant Group)
Offene Gewichte
— — 262K Klein OpenRouter, Vercel AI Gateway, Nous Portal, UnoRouter
52
Laguna XS 2.1
Poolside
ReasoningCoding
— — 262K Klein OpenRouter, NVIDIA Build, Nous Portal, UnoRouter
53
DiffusionGemma 26B
Google DeepMind · 26B · 4B
Tool-NutzungOffene Gewichte
9.5 67% 262K Mittelgroß UnoRouter
54 — — 262K Mittelgroß Together AI
55
Holo3 35B A3B
35B · 3B
— — 262K Mittelgroß Together AI
56 — — 262K Mittelgroß UnoRouter
57
Qwen SEA-LION V4.5 27B
Alibaba Cloud · 27B
Tool-NutzungOffene Gewichte
— — 262K Mittelgroß UnoRouter
58
Nvidia Nemotron 3 Nano
NVIDIA · 30B · 3B
ReasoningTool-NutzungOffene Gewichte
8.9 76% 262K Mittelgroß Together AI
59
Qwen3 4B
Alibaba Cloud · 4B
ReasoningOffene Gewichte
8.8 67% 262K Klein Together AI
60
Qwen 3.5 2B
Alibaba Cloud · 2B
ReasoningCodingOffene Gewichte
6.9 46% 262K Klein Together AI
61 — — 262K Klein UnoRouter
62
Mistral Large 3 675B
Mistral AI · 675B
CodingTool-NutzungOffene Gewichte
— — 256K Sehr groß UnoRouter
63
Mistral Small
Mistral AI · 24B
Tool-NutzungOffene Gewichte
5.8 38% 256K Mittelgroß Mistral AI, UnoRouter
64
Codestral Mamba
Mistral AI
Coding
— — 256K Klein Mistral AI
65
Codestral
Mistral AI
Coding
— — 256K Klein UnoRouter

Benchmarks: Artificial Analysis (24. September 2026). Ein Strich bedeutet, dass das Modell nicht in deren Rangliste steht; wir schätzen Werte nie selbst.

FAQ

Welches ist das beste kostenlose Modell für langer kontext?

GLM 5.3 führt derzeit — kostenlos bei Together AI, NVIDIA Build. Danach folgen Mimo V2.6 Pro, Kimi K3, GLM 5.3 Flash.

Siehe auch