Ling 3.1 Flash
Fast and lightToolsinclusionAI · Released October 2026
Ling 3.1 Flash is a hybrid reasoning mixture-of-experts model from inclusionAI, with 25B active parameters out of 560B total.
It is part of inclusionAI's open-weight Ling line, where Flash leads on speed, Flash VL adds vision, and the Sante and Fin variants are tuned for health and finance.
Facts and pricing
- Provider
- inclusionAI
- Context window
- 262K tokens
- Input price
- 0 credits / 1M tokens ($0 raw)
- Output price
- 0 credits / 1M tokens ($0 raw)
- Vision (image input)
- No
- Tool calling
- Yes
- Extended reasoning
- No
Credits are what Zeplik bills: 1 credit = $0.10, computed from the raw provider rate with a 1.20x margin. Raw prices shown per 1M tokens.
What Ling 3.1 Flash is best for
- High-volume everyday tasks where speed and cost matter: summaries, drafts, quick questions
- Long documents: the 262K-token window holds lengthy reports, contracts or papers whole
- Tool use and agents: reliably calls functions, so it can search, run skills and drive workflows
Example prompts
Prompts that suit a fast and light model like Ling 3.1 Flash:
- Summarize this article in five bullet points
- Rewrite this email to be shorter and friendlier
- Give me ten name ideas for a hiking newsletter
- Translate this paragraph to Spanish and keep the tone
inclusionAI family
| Model | Context | Input cr/M | Output cr/M | Released |
|---|---|---|---|---|
| Ling 3.1 Flash | 262K | 0 | 0 | October 2026 |
| Ling 3.0 Flash VL | 262K | 0.25 | 0.74 | September 2026 |
| Ling 3.0 Flash Sante (free) | 262K | 0 | 0 | September 2026 |
| Ling 3.0 Flash Fin | 262K | 0.50 | 1.5 | August 2026 |
| Ling 3.0 Flash | 262K | 0.25 | 0.76 | July 2026 |
Frequently asked questions
- What is Ling 3.1 Flash?
- Ling 3.1 Flash is an AI model by inclusionAI. Ling 3.1 Flash is a hybrid reasoning mixture-of-experts model from inclusionAI, with 25B active parameters out of 560B total.
- How much does Ling 3.1 Flash cost on Zeplik?
- Input tokens cost 0 credits per million and output tokens 0 credits per million (1 credit = $0.10; the raw provider rates are $0 and $0 per million). New accounts start with free credits.
- How long can a conversation with Ling 3.1 Flash be?
- Ling 3.1 Flash has a 262K-token context window (262,144 tokens), which covers the conversation plus any documents you attach.
- Does Ling 3.1 Flash support images and tools?
- Ling 3.1 Flash is text-only and supports tool calling.
Related models
Ling 3.0 Flash Sante is a health and medicine-focused mixture-of-experts model from InclusionAI, built on Ling 3.0 Flash with 5.1B active parameters out of 124B total.
Ling 3.0 Flash Fin is a finance-focused mixture-of-experts model from InclusionAI, built on Ling 3.0 Flash with 5.1B active parameters out of 124B total.
*Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (MoE) model*, with approximately *5.1B parameters activated per token*.
Claude Haiku 5.5 is Anthropic's small, fast model for high-volume, cost-sensitive work such as summarization, subagents, and browser use.
Nano Banana 2.1 (Gemini Nano Banana 2.1) is Google's image generation and editing model on the Flash tier, succeeding Nano Banana 2 and Nano Banana Pro.
More you can do with Ling 3.1 Flash
- 467 AI skillsReady-to-run expert methods for writing, data, code and more, each running on Ling 3.1 Flash or any model you pick.
- 997 integrationsConnect Gmail, Slack, GitHub, Notion and hundreds more so the assistant can work in your apps.
- Simple pricingPlus, Pro and Max plans with a monthly credit allowance that works across every model.