Skip to main content

Ling 3.0 Flash

Fast and lightTools

inclusionAI · Released July 2026

*Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (MoE) model*, with approximately *5.1B parameters activated per token*. The model is designed with *token efficiency and production-scale agentic inference* as key priorities, enabling developers...

Facts and pricing

Provider
inclusionAI
Context window
262K tokens
Input price
0.23 credits / 1M tokens ($0.021 raw)
Output price
0.69 credits / 1M tokens ($0.063 raw)
Vision (image input)
No
Tool calling
Yes
Extended reasoning
No

Credits are what Zeplik bills: 1 credit = $0.10, computed from the raw provider rate with a 1.10x margin. Raw prices shown per 1M tokens.

Try Ling 3.0 Flash now

Ask Ling 3.0 Flash anything. Your prompt opens in the Zeplik app with this model selected.

Ling 3.0 Flash

Or open Zeplik with Ling 3.0 Flash already selected

What Ling 3.0 Flash is best for

inclusionAI family

ModelContextInput cr/MOutput cr/MReleased
Ling 3.0 Flash Sante (free)262K00September 2026
Ling 3.0 Flash Fin (free)262K00August 2026
Ling 3.0 Flash Fin262K0.662.0August 2026
Ling 3.0 Flash262K0.230.69July 2026

Frequently asked questions

What is Ling 3.0 Flash?
Ling 3.0 Flash is an AI model by inclusionAI. *Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (MoE) model*, with approximately *5.1B parameters activated per token*. The model is designed with *token efficiency and production-scale agentic inference* as key priorities, enabling developers...
How much does Ling 3.0 Flash cost on Zeplik?
Input tokens cost 0.23 credits per million and output tokens 0.69 credits per million (1 credit = $0.10; the raw provider rates are $0.021 and $0.063 per million). New accounts start with free credits.
How long can a conversation with Ling 3.0 Flash be?
Ling 3.0 Flash has a 262K-token context window (262,144 tokens), which covers the conversation plus any documents you attach.
Does Ling 3.0 Flash support images and tools?
Ling 3.0 Flash is text-only and supports tool calling.

Related models

Ling 3.0 Flash Sante (free)Fast and light

Ling 3.0 Flash Sante is a health and medicine-focused mixture-of-experts model from InclusionAI, built on Ling 3.0 Flash with 5.1B active parameters out of 124B total. It is designed for...

262K ctx0 cr/M outSeptember 2026
Ling 3.0 Flash Fin (free)Fast and light

Ling 3.0 Flash Fin is a finance-focused mixture-of-experts model from InclusionAI, built on Ling 3.0 Flash with 5.1B active parameters out of 124B total. It is designed for real-world investment...

262K ctx0 cr/M outAugust 2026
Ling 3.0 Flash FinFast and light

Ling 3.0 Flash Fin is a finance-focused mixture-of-experts model from InclusionAI, built on Ling 3.0 Flash with 5.1B active parameters out of 124B total. It is designed for real-world investment...

262K ctx2.0 cr/M outAugust 2026
GPT-6 AstraFrontier

GPT-6 Astra is OpenAI's flagship model for demanding end-to-end work. It is suited for advanced analysis, software engineering, deep research, scientific work, and document creation, with particular strengths in long-horizon...

1.1M ctx550 cr/M outSeptember 2026
GPT-6 Astra ProFrontier

GPT-6 Astra Pro is the same underlying model as [GPT-6 Astra](https://openrouter.ai/openai/gpt-6-astra), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode

1.1M ctx550 cr/M outSeptember 2026
Gemini 3.8 Flash (batch)Fast and light

Gemini 3.8 Flash is Google's most intelligent Flash model with significant gains from 3.7 Flash across software engineering, agentic tasks, and multi-step reasoning.

1.0M ctx20.6 cr/M outSeptember 2026

More you can do with Ling 3.0 Flash

Ling 3.0 Flash - inclusionAI AI Model | Zeplik Chat