Qwen3.8 2.4T A95B
FrontierToolsQwen · Released August 2026
Qwen3.8 2.4T A95B is an open-weight sparse mixture-of-experts model from Qwen and the open-weight variant of [Qwen3.8 Max](/qwen/qwen3.8-max), with 95 billion active parameters out of 2.4 trillion total. It is...
It is part of Alibaba's Qwen family, a broad catalog where Max leads on capability, Plus balances cost, and Flash and the open-weight sizes cover high-volume work.
Facts and pricing
- Provider
- Qwen
- Context window
- 1.0M tokens
- Input price
- 22.0 credits / 1M tokens ($2.00 raw)
- Output price
- 66.0 credits / 1M tokens ($6.00 raw)
- Vision (image input)
- No
- Tool calling
- Yes
- Extended reasoning
- No
Credits are what Zeplik bills: 1 credit = $0.10, computed from the raw provider rate with a 1.10x margin. Raw prices shown per 1M tokens.
Try Qwen3.8 2.4T A95B now
What Qwen3.8 2.4T A95B is best for
- Hard problems where answer quality matters more than cost: strategy, analysis, difficult writing
- Very long documents and codebases: a 1.0M-token window fits entire books or repositories in one conversation
- Tool use and agents: reliably calls functions, so it can search, run skills and drive workflows
Example prompts
Prompts that suit a frontier model like Qwen3.8 2.4T A95B:
- Review this business plan and give me the three hardest questions an investor would ask
- Rewrite this announcement so it lands with both engineers and executives
- I am choosing between two job offers, help me think through the tradeoffs
- Draft a technical design doc for a rate limiter, including failure modes
Qwen family
| Model | Context | Input cr/M | Output cr/M | Released |
|---|---|---|---|---|
| Qwen3.8 2.4T A95B | 1.0M | 22.0 | 66.0 | August 2026 |
| Qwen3.8 Max | 1M | 22.0 | 66.0 | August 2026 |
| Qwen3.7 Flash | 1M | 0.33 | 1.4 | July 2026 |
| Qwen3.7 Plus | 1M | 3.5 | 14.1 | June 2026 |
| Qwen3.7 Max | 1M | 16.2 | 48.7 | May 2026 |
| Qwen3.5 Plus 2026-04-20 | 1M | 3.3 | 19.8 | April 2026 |
| Qwen3.6 Flash | 1M | 2.1 | 12.4 | April 2026 |
| Qwen3.6 35B A3B | 262K | 1.5 | 11.0 | April 2026 |
| Qwen3.6 Max Preview | 262K | 11.3 | 67.8 | April 2026 |
| Qwen3.6 27B | 262K | 6.6 | 39.6 | April 2026 |
Frequently asked questions
- What is Qwen3.8 2.4T A95B?
- Qwen3.8 2.4T A95B is an AI model by Qwen. Qwen3.8 2.4T A95B is an open-weight sparse mixture-of-experts model from Qwen and the open-weight variant of [Qwen3.8 Max](/qwen/qwen3.8-max), with 95 billion active parameters out of 2.4 trillion total. It is...
- How much does Qwen3.8 2.4T A95B cost on Zeplik?
- Input tokens cost 22.0 credits per million and output tokens 66.0 credits per million (1 credit = $0.10; the raw provider rates are $2.00 and $6.00 per million). New accounts start with free credits.
- How long can a conversation with Qwen3.8 2.4T A95B be?
- Qwen3.8 2.4T A95B has a 1.0M-token context window (1,048,576 tokens), which covers the conversation plus any documents you attach.
- Does Qwen3.8 2.4T A95B support images and tools?
- Qwen3.8 2.4T A95B is text-only and supports tool calling.
Related models
Qwen3.8 Max is the flagship model in Alibaba's Qwen3.8 series, the general-availability successor to the Qwen3.8 Max Preview. It is a multimodal reasoning model intended for complex reasoning, visual understanding,...
Qwen3.7 Flash is a vision-language reasoning model from Alibaba. It is suited for multimodal agents, visual coding, search, and computer interaction, with strengths in object recognition, spatial understanding, and real-world...
Qwen3.7-Plus is a cost-effective model in Alibaba's Qwen3.7 series. It supports text and image input with text output, building on the series' text capabilities with a comprehensive upgrade to its...
Qwen3.7-Max is the flagship model in Alibaba's Qwen3.7 series. It supports text input and output and is designed for agent-centric workloads, with particular strengths in coding, office and productivity tasks,...
Qwen3.5 Plus (April 2026) is a large-scale multimodal language model from Alibaba. It accepts text, image, and video input and produces text output, with a 1M token context window. This...
Qwen3.6 Flash is a fast, efficient language model from Alibaba's Qwen 3.6 series. It supports text, image, and video input with a 1M token context window. Tiered pricing kicks in...
More you can do with Qwen3.8 2.4T A95B
- 465 AI skillsReady-to-run expert methods for writing, data, code and more, each running on Qwen3.8 2.4T A95B or any model you pick.
- 997 integrationsConnect Gmail, Slack, GitHub, Notion and hundreds more so the assistant can work in your apps.
- Simple pricingPlus, Pro and Max plans with a monthly credit allowance that works across every model.