Ling 3.1 Flash
Ling 3.1 Flash is a language model from inclusionAI (Ant Group). It is currently free at 1 provider with up to 262K tokens of context. Its weights are open, so it can also run locally.
- Developer
- inclusionAI (Ant Group)
- Context
- 262K tokens
- Type
- Text
- License
- Open weights
- Free since
- September 29, 2026
Where to use Ling 3.1 Flash for free
Each row is a separate free endpoint with its own limits.
| Provider | Context | Limit | Card | |
|---|---|---|---|---|
Vercel AI Gateway inclusionai/ling-3.1-flash |
262K | A valid payment card must be on file even for zero-price models. Free promotions may change; availability is refreshed from the Models API. | Yes | Open |
Vercel AI Gateway inclusionai/ling-3.1-flash-free |
262K | A valid payment card must be on file even for zero-price models. Free promotions may change; availability is refreshed from the Models API. | Yes | Open |
Quick start
curl https://ai-gateway.vercel.sh/v1/chat/completions \
-H "Authorization: Bearer $VERCEL_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "inclusionai/ling-3.1-flash",
"messages": [{"role": "user", "content": "Hello!"}]
}'from openai import OpenAI
client = OpenAI(
base_url="https://ai-gateway.vercel.sh/v1",
api_key="YOUR_VERCEL_API_KEY",
)
reply = client.chat.completions.create(
model="inclusionai/ling-3.1-flash",
messages=[{"role": "user", "content": "Hello!"}],
)
print(reply.choices[0].message.content)import OpenAI from "openai";
const client = new OpenAI({
baseURL: "https://ai-gateway.vercel.sh/v1",
apiKey: process.env.VERCEL_API_KEY,
});
const reply = await client.chat.completions.create({
model: "inclusionai/ling-3.1-flash",
messages: [{ role: "user", content: "Hello!" }],
});
console.log(reply.choices[0].message.content);Example for Vercel AI Gateway. Any OpenAI-compatible client works: Cursor, Cline, OpenCode, Open WebUI.
Benchmarks
This model is not on the Artificial Analysis leaderboard yet. We do not publish our own estimates.
Similar free models
FAQ
Is Ling 3.1 Flash free?
Yes. As of September 29, 2026, Ling 3.1 Flash is free at Vercel AI Gateway. Each provider has its own limits — see the table above.
What is the context window of Ling 3.1 Flash?
The largest context among free routes is 262,144 tokens. Some providers cap it lower.
How do I call Ling 3.1 Flash via API?
Use any OpenAI-compatible client with base URL https://ai-gateway.vercel.sh/v1, model ID inclusionai/ling-3.1-flash, and a Vercel AI Gateway key.
Is Ling 3.1 Flash open-weight?
Yes, the weights are published — you can download and run it locally, for example with Ollama or LM Studio.