LLM Perks

API gratuites de modèles open-weight

Des modèles open-weight que vous pouvez appeler gratuitement via l'API et, si besoin, exécuter sur votre propre matériel sans dépendance à un fournisseur.

125 modèles11 fournisseursjusqu'à 10.5M contextemeilleur : GLM 5.3Mis à jour:
En savoir plus

Open-weight signifie que les paramètres sont publiés, mais les licences diffèrent : Llama et Gemma ont leurs propres conditions, tandis que DeepSeek, Qwen et gpt-oss sont plus permissifs. Vérifiez la licence avant tout usage commercial.

Taille
125 trouvé(s)
# Modèle ↕ Indice AA ↕ GPQA ↕ Contexte ↕ Taille ↕ Gratuit chez ↕
1
GLM 5.3
Z.ai (Zhipu AI)
RaisonnementCodeAppel d'outilsPoids ouverts
44.8 92% 1M Très grand Together AI, NVIDIA Build
2
Kimi K3
Moonshot AI
VisionRaisonnementCodeAppel d'outilsPoids ouverts
43.6 94% 1M Très grand NVIDIA Build, UnoRouter
3
GLM 5.3 Flash
Z.ai (Zhipu AI)
RaisonnementCodeAppel d'outilsPoids ouverts
41.8 91% 1M Taille moyenne OrcaRouter, NVIDIA Build, UnoRouter
4
DeepSeek V4.1 Flash
DeepSeek
RaisonnementCodeAppel d'outilsPoids ouverts
39.5 — 1M Très grand NVIDIA Build, UnoRouter
5
Qwen3.8 27B
Alibaba Cloud · 27B
RaisonnementCodeAppel d'outilsPoids ouverts
33.7 91% 262K Taille moyenne OpenRouter, Groq, Cloudflare Workers AI, UnoRouter
6
DeepSeek V4 Pro
DeepSeek
RaisonnementCodeAppel d'outilsPoids ouverts
36.0 93% 1M Très grand UnoRouter
7
GLM 5.2
Z.ai (Zhipu AI)
RaisonnementCodeAppel d'outilsPoids ouverts
33.7 90% 1M Très grand Together AI
8
DeepSeek V4 Flash
DeepSeek
RaisonnementCodeAppel d'outilsPoids ouverts
34.3 91% — Très grand OrcaRouter
9
Inkling Small
Thinking Machines Lab
RaisonnementPoids ouverts
27.8 90% 1M Petit OpenRouter
10
Qwen 3.6 Plus
Alibaba Cloud · New
RaisonnementCodeAppel d'outilsPoids ouverts
27.0 88% 1M Très grand UnoRouter
11
Nemotron 3 Ultra
NVIDIA · 550B · 55B
RaisonnementCodeAppel d'outilsPoids ouverts
22.9 87% 1M Très grand OpenRouter, NVIDIA Build, UnoRouter
12
Inkling
Thinking Machines Lab
RaisonnementPoids ouverts
25.0 87% 1M Petit OpenRouter
13
Kimi K2.6
Moonshot AI
RaisonnementCodeAppel d'outilsPoids ouverts
27.0 91% — Très grand NVIDIA Build
14
Hy3
Tencent
Poids ouverts
25.3 90% — Petit OrcaRouter
15
Ling 3.0 Flash Fin
inclusionAI (Ant Group)
Poids ouverts
22.6 — 262K Petit Nous Portal
16
Gemma 4 31B
Google DeepMind · 31B
VisionAppel d'outilsPoids ouverts
19.0 86% 262K Taille moyenne OpenRouter, Together AI, NVIDIA Build
17
Gemma 4 26B A4B
Google DeepMind · 26B · 4B
VisionAppel d'outilsPoids ouverts
16.7 79% 262K Taille moyenne OpenRouter, Together AI, Cloudflare Workers AI, UnoRouter
18
GLM 4.7
Z.ai (Zhipu AI)
RaisonnementCodeAppel d'outilsPoids ouverts
22.2 86% 203K Très grand Together AI
19
LongCat 2.0
Meituan
RaisonnementAppel d'outilsPoids ouverts
19.1 78% 1M Très grand Nous Portal
20
Step 3.7 Flash
StepFun
RaisonnementPoids ouverts
19.5 81% 262K Petit Nous Portal
21
Qwen 3.5 35B
Alibaba Cloud · 35B · 3B
RaisonnementCodeAppel d'outilsPoids ouverts
19.3 85% 262K Taille moyenne Together AI
22
Qwen 3.5 122B
Alibaba Cloud · 122B · 10B
RaisonnementCodeAppel d'outilsPoids ouverts
17.7 83% 262K Très grand Together AI, UnoRouter
23
Qwen 3.6 35B
Alibaba Cloud · 35B · 3B
RaisonnementCodeAppel d'outilsPoids ouverts
18.2 84% 262K Taille moyenne Together AI
24
MiniMax M2
MiniMax · 230B · 10B
RaisonnementCodeAppel d'outilsPoids ouverts
18.6 78% 197K Très grand Together AI
25
Nemotron 3.5 Lightning
NVIDIA · 30B · 3B
RaisonnementAppel d'outilsPoids ouverts
12.9 74% 1M Taille moyenne OpenRouter, NVIDIA Build, UnoRouter
26
Nemotron 3 Super
NVIDIA · 120B · 12B
RaisonnementCodeAppel d'outilsPoids ouverts
12.8 80% 262K Très grand OpenRouter, NVIDIA Build, UnoRouter
27
GLM 4.7 Flash
Z.ai (Zhipu AI)
RaisonnementCodeAppel d'outilsPoids ouverts
14.9 58% 131K Taille moyenne Cloudflare Workers AI, UnoRouter
28
Llama 4 Maverick Instruct Fp4
Meta · 400B · 17B
VisionAppel d'outilsPoids ouverts
— — 1M Très grand Together AI
29
LongCat 2.5 Preview
Meituan · New
RaisonnementAppel d'outilsPoids ouverts
— — 1M Très grand Nous Portal
30
Nemotron 3 Nano Omni
NVIDIA · 30B · 3B
VisionRaisonnementAppel d'outilsPoids ouverts
10.3 47% 262K Taille moyenne +1OpenRouter, Together AI, NVIDIA Build, TokenRouter, UnoRouter
31
Gemma 4 12B
Google DeepMind · 12B
VisionAppel d'outilsPoids ouverts
14.2 75% 262K Taille moyenne Together AI
32
Nemotron 3 120B
NVIDIA · 120B · 12B · New
RaisonnementAppel d'outilsPoids ouverts
— — 256K Très grand Cloudflare Workers AI
33
Mistral Large 3 675B
Mistral AI · 675B
CodeAppel d'outilsPoids ouverts
— — 256K Très grand UnoRouter
34
Qwen 3.5 9B
Alibaba Cloud · 9B
RaisonnementCodePoids ouverts
13.7 81% 262K Petit Together AI
35
gpt-oss-120b
OpenAI · 117B · 5.1B
RaisonnementCodeAppel d'outilsPoids ouverts
11.6 78% 131K Très grand Groq, Cloudflare Workers AI, UnoRouter
36
gpt-oss-20b
OpenAI · 21B · 3.6B
RaisonnementAppel d'outilsPoids ouverts
10.0 61% 131K Taille moyenne Groq, Cloudflare Workers AI, NVIDIA Build, UnoRouter
37
Qwen3 VL 235B
Alibaba Cloud · 235B · 22B
VisionRaisonnementAppel d'outilsPoids ouverts
13.4 77% 262K Très grand Together AI
38
Qwen 3.5 4B
Alibaba Cloud · 4B
RaisonnementCodePoids ouverts
13.1 77% 262K Petit UnoRouter
39
Cogito V1 Preview Llama 70B
Meta · 70B
Appel d'outilsPoids ouverts
— — 131K Très grand Together AI
40
Meta Llama 3.1 70B
Meta · 70B
Appel d'outilsPoids ouverts
— — 131K Très grand Together AI
41
GLM 4.7 Fp4
Z.ai (Zhipu AI)
RaisonnementCodeAppel d'outilsPoids ouverts
— — 203K Très grand Together AI
42
GLM 5 Fp4
Z.ai (Zhipu AI)
RaisonnementCodeAppel d'outilsPoids ouverts
— — 203K Très grand Together AI
43
MiniMax M2.5 Fp4
MiniMax
RaisonnementCodeAppel d'outilsPoids ouverts
— — 8K Très grand Together AI
44
GLM Ocr
Z.ai (Zhipu AI)
Appel d'outilsPoids ouverts
— — 131K Très grand Together AI
45
Llama 3.3 70B Instruct Fp8 Fast
Meta · 70B · New
Appel d'outilsPoids ouverts
— — 24K Très grand Cloudflare Workers AI
46
Nvidia Nemotron 3 Super
NVIDIA · 120B · 12B
RaisonnementCodeAppel d'outilsPoids ouverts
12.8 80% 262K Très grand Together AI
47
MiniMax M1 80K
MiniMax
RaisonnementAppel d'outilsPoids ouverts
11.7 70% 1M Très grand Together AI
48
Qwen3 235B
Alibaba Cloud · 235B · 22B
RaisonnementCodeAppel d'outilsPoids ouverts
12.7 79% 262K Très grand Together AI
49
Qwen3 Next 80B
Alibaba Cloud · 80B · 3B
RaisonnementCodeAppel d'outilsPoids ouverts
11.2 76% 131K Très grand Together AI, UnoRouter
50
North Mini Code
Cohere
CodePoids ouverts
9.9 76% 262K Petit OpenRouter, UnoRouter
51
Qwen3 30B
Alibaba Cloud · 30B · 3B
RaisonnementAppel d'outilsPoids ouverts
9.8 71% 262K Taille moyenne Together AI, Cloudflare Workers AI
52
Mistral Nemo
Mistral AI · 12B
Appel d'outilsPoids ouverts
— — 131K Taille moyenne Mistral AI, Together AI, NVIDIA Build
53
MiniMax M1 40K
MiniMax
RaisonnementAppel d'outilsPoids ouverts
10.0 68% 1M Très grand Together AI
54
Llama 4 Scout
Meta · 109B · 17B
VisionAppel d'outilsPoids ouverts
8.1 59% 10.5M Très grand Together AI, Cloudflare Workers AI
55
QwQ 32B
Alibaba Cloud · 32B
RaisonnementAppel d'outilsPoids ouverts
9.5 59% 131K Taille moyenne Cloudflare Workers AI, UnoRouter
56
GLM 5.3 Flash Search
Z.ai (Zhipu AI)
RaisonnementCodeAppel d'outilsPoids ouverts
— — 1M Taille moyenne UnoRouter
57
Qwen3 Coder 30B
Alibaba Cloud · 30B · 3B
RaisonnementCodeAppel d'outilsPoids ouverts
9.6 52% 262K Très grand Together AI
58
Ling 3.0 Flash Sante
inclusionAI (Ant Group)
Poids ouverts
— — 262K Petit OpenRouter, Vercel AI Gateway, Nous Portal, UnoRouter
59
DiffusionGemma 26B
Google DeepMind · 26B · 4B
Appel d'outilsPoids ouverts
9.5 67% 262K Taille moyenne UnoRouter
60
Llama 4 Maverick
Meta · 400B · 17B
VisionAppel d'outilsPoids ouverts
10.0 67% 131K Très grand UnoRouter
61
Qwen SEA-LION V4.5 27B
Alibaba Cloud · 27B
Appel d'outilsPoids ouverts
— — 262K Taille moyenne UnoRouter
62
LFM2.5-2.6B
Liquid AI · 2.6B
Poids ouverts
8.4 56% 128K Petit OpenRouter, UnoRouter
63
Llama 3.2 11B Vision
Meta · 11B
VisionAppel d'outilsPoids ouverts
5.4 22% 131K Taille moyenne Together AI, Cloudflare Workers AI, NVIDIA Build, UnoRouter
64
Nvidia Nemotron 3 Nano
NVIDIA · 30B · 3B
RaisonnementAppel d'outilsPoids ouverts
8.9 76% 262K Taille moyenne Together AI
65
DeepSeek R1 Distill Qwen 32B
DeepSeek · 32B
RaisonnementAppel d'outilsPoids ouverts
8.4 62% 80K Taille moyenne Cloudflare Workers AI, UnoRouter
66
Qwen3 4B
Alibaba Cloud · 4B
RaisonnementPoids ouverts
8.8 67% 262K Petit Together AI
67
Mixtral 8x7b Instruct V01
Mistral AI · 46.7B · 12.9B
Appel d'outilsPoids ouverts
— — 16K Taille moyenne Together AI
68
Cogito V1 Preview Qwen 32B
Alibaba Cloud · 32B
Appel d'outilsPoids ouverts
— — 131K Taille moyenne Together AI
69
Cogito V1 Preview Qwen 14B
Alibaba Cloud · 14B
Appel d'outilsPoids ouverts
— — 131K Taille moyenne Together AI
70
Qwen2.5 14B
Alibaba Cloud · 14B
Appel d'outilsPoids ouverts
— — 131K Taille moyenne Together AI
71
Medgemma 27B Text
Google DeepMind · 27B
Appel d'outilsPoids ouverts
— — 131K Taille moyenne Together AI
72
DeepSeek Ocr 2
DeepSeek
Appel d'outilsPoids ouverts
— — 8K Taille moyenne Together AI
73
Gemma SEA-LION V4 27B
Google DeepMind · 27B · New
Appel d'outilsPoids ouverts
— — 128K Taille moyenne Cloudflare Workers AI
74
Mistral Small 3.1 24B
Mistral AI · 24B · New
Appel d'outilsPoids ouverts
— — 128K Taille moyenne Cloudflare Workers AI
75
Mistral Nemotron
Mistral AI · 12B
Appel d'outilsPoids ouverts
— — 128K Taille moyenne UnoRouter
76
Qwen SEA-LION V4 32B
Alibaba Cloud · 32B
Appel d'outilsPoids ouverts
— — 33K Taille moyenne UnoRouter
77
Qwen3
Alibaba Cloud
RaisonnementAppel d'outilsPoids ouverts
— — 25K Taille moyenne UnoRouter
78
Llama 3.3 Nemotron Super 49B
NVIDIA · 49B
Appel d'outilsPoids ouverts
8.9 64% 16K Taille moyenne Together AI
79
Gemma 4 E4b
Google DeepMind
VisionPoids ouverts
8.9 58% 131K Petit Together AI
80
Nemotron 3 Nano
NVIDIA · 30B · 3B
Appel d'outilsPoids ouverts
8.9 76% — Taille moyenne NVIDIA Build
81
Devstral Small
Mistral AI
CodeAppel d'outilsPoids ouverts
8.7 43% 131K Taille moyenne Together AI
82
Magistral Small
Mistral AI
RaisonnementAppel d'outilsPoids ouverts
8.6 66% 41K Taille moyenne Together AI
83
Qwen3 32B
Alibaba Cloud · 32B
RaisonnementAppel d'outilsPoids ouverts
8.6 67% 41K Taille moyenne Together AI
84
Llama 3.1 8B
Meta · 8B
Poids ouverts
6.9 26% 32K Petit Together AI, Cloudflare Workers AI
85
Mistral Small
Mistral AI · 24B
Appel d'outilsPoids ouverts
5.8 38% 256K Taille moyenne Mistral AI, UnoRouter
86
Qwen3 14B
Alibaba Cloud · 14B
RaisonnementAppel d'outilsPoids ouverts
8.2 60% 2K Taille moyenne Together AI
87
Qwen2.5 Coder 32B
Alibaba Cloud · 32B
CodeAppel d'outilsPoids ouverts
6.7 42% 131K Taille moyenne Cloudflare Workers AI, UnoRouter
88
Llama 3.2 90B Vision
Meta · 90B
VisionAppel d'outilsPoids ouverts
6.4 43% 16K Très grand Together AI, NVIDIA Build
89
Qwen 3.5 2B
Alibaba Cloud · 2B
RaisonnementCodePoids ouverts
6.9 46% 262K Petit Together AI
90
Gemma 4 E2b
Google DeepMind
VisionPoids ouverts
7.8 43% 131K Petit Together AI
91
Llama 3.3 70B
Meta · 70B
Appel d'outilsPoids ouverts
7.7 50% 131K Très grand Together AI
92
Qwen2.5 72B
Alibaba Cloud · 72B
Appel d'outilsPoids ouverts
7.7 49% 131K Très grand Together AI
93
GLM 4 5v
Z.ai (Zhipu AI)
RaisonnementCodeAppel d'outilsPoids ouverts
7.6 68% 66K Très grand Together AI
94
Llama 3.1 Nemotron Ultra 253B
NVIDIA · 253B
Appel d'outilsPoids ouverts
7.5 73% — Très grand NVIDIA Build
95
Nemotron Nano 12B V2 VL
NVIDIA · 12B
VisionAppel d'outilsPoids ouverts
7.5 57% 131K Taille moyenne UnoRouter
96
Nemotron Nano 9B V2
NVIDIA · 9B
Poids ouverts
7.4 57% 131K Petit UnoRouter
97
Llama 3.1 405B
Meta · 405B
Appel d'outilsPoids ouverts
7.3 52% 131K Très grand Together AI
98
Qwen3 8B
Alibaba Cloud · 8B
RaisonnementPoids ouverts
7.3 59% 41K Petit Together AI
99
Qwen2.5 32B
Alibaba Cloud · 32B
Appel d'outilsPoids ouverts
6.9 47% 131K Taille moyenne Together AI
100
Llama 3.1 70B
Meta · 70B
Appel d'outilsPoids ouverts
6.6 41% 16K Très grand Together AI
101
Gemma 3 4B
Google DeepMind · 4B
VisionPoids ouverts
4.8 29% 66K Petit Together AI, NVIDIA Build
102
Llama 3.2 1B
Meta · 1B
Poids ouverts
4.8 20% 131K Petit Together AI, Cloudflare Workers AI
103
Ling 3.1 Flash
inclusionAI (Ant Group) · New
Poids ouverts
— — 262K Petit Vercel AI Gateway
104
Llama 3.2 3B
Meta · 3B · New
Poids ouverts
5.7 26% 80K Petit Cloudflare Workers AI
105
Molmo 7B D 0924
7B
Poids ouverts
5.6 24% 4K Petit Together AI
106
Sarvam M
Poids ouverts
5.3 42% 33K Petit Together AI
107
Qwen 3.1 7B
Alibaba Cloud · 1.7B
RaisonnementPoids ouverts
5.2 36% 41K Petit Together AI
108
Gemma 3 270M
Google DeepMind · 270M
Poids ouverts
5.1 22% 33K Petit Together AI
109
Granite 4.0 Micro
IBM
Poids ouverts
5.1 34% 131K Petit UnoRouter
110
DeepSeek R1 Distill Qwen 7B
DeepSeek · 7B
RaisonnementPoids ouverts
— — 131K Petit Together AI
111
Cogito V1 Preview Llama 8B
Meta · 8B
Poids ouverts
— — 131K Petit Together AI
112
Qwen2.5 3B
Alibaba Cloud · 3B
Poids ouverts
— — 33K Petit Together AI
113
Qwen2.5 1 5B
Alibaba Cloud · 1.5B
Poids ouverts
— — 131K Petit Together AI
114
Qwen2.5 7B
Alibaba Cloud · 7B
Poids ouverts
— — 131K Petit Together AI
115
Meta Llama 3.1 8B
Meta · 8B
Poids ouverts
— — 131K Petit Together AI
116
Granite 4.0 H Micro
IBM · New
Poids ouverts
— — 131K Petit Cloudflare Workers AI
117
Gemma 7B
Google DeepMind · 7B · New
Poids ouverts
— — 4K Petit Cloudflare Workers AI
118
Mistral Nemo Minitron 8B 8K
Mistral AI · 8B
Poids ouverts
— — 8K Petit NVIDIA Build
119
Qwen2.5 VL 7B
Alibaba Cloud · 7B
VisionPoids ouverts
— — 131K Petit UnoRouter
120
Gemma 3 27B
Google DeepMind · 27B
VisionAppel d'outilsPoids ouverts
4.9 43% 66K Taille moyenne Together AI
121
Gemma 3 1B
Google DeepMind · 1B
Poids ouverts
4.8 24% 33K Petit Together AI
122
Qwen 3.0 6B
Alibaba Cloud · 600M
Poids ouverts
4.8 23% 41K Petit Together AI
123
Gemma 3 12B
Google DeepMind · 12B
VisionAppel d'outilsPoids ouverts
3.8 35% — Taille moyenne NVIDIA Build
124
Nemotron 3.5 Content Safety
NVIDIA · Modération
Poids ouverts
— — 128K Taille moyenne OpenRouter
125
gpt-oss-safeguard-20b
OpenAI · 21B · 3.6B · Modération
Poids ouverts
— — 131K Taille moyenne UnoRouter

Benchmarks : Artificial Analysis (30 septembre 2026). Un tiret signifie que le modèle ne figure pas dans leur classement ; nous n'estimons jamais les scores nous-mêmes.

FAQ

Quels modèles open-weight sont gratuits en ce moment ?

125 modèles sont gratuits en ce moment. En tête selon l'indice Artificial Analysis : GLM 5.3, Kimi K3, GLM 5.3 Flash, DeepSeek V4.1 Flash, Qwen3.8 27B, DeepSeek V4 Pro.

Où utiliser GLM 5.3 gratuitement ?

GLM 5.3 est gratuit chez Together AI, NVIDIA Build. Chaque fournisseur est compatible OpenAI : il suffit de changer la base URL et la clé.

Quelles sont les limites de l'accès gratuit ?

Open-weight signifie que les paramètres sont publiés, mais les licences diffèrent : Llama et Gemma ont leurs propres conditions, tandis que DeepSeek, Qwen et gpt-oss sont plus permissifs. Vérifiez la licence avant tout usage commercial.

Voir aussi