Magnum v4 72B
FrontierAnthracite · Released October 2024 · Knowledge cutoff June 2024
This is a series of models designed to replicate the prose quality of the Claude 3 models, specifically Sonnet(https://openrouter.ai/anthropic/claude-3.5-sonnet) and Opus(https://openrouter.ai/anthropic/claude-3-opus). The model is fine-tuned on top of [Qwen2.5 72B](https://openrouter.ai/qwen/qwen-2.5-72b-instruct).
Facts and pricing
- Provider
- Anthracite
- Context window
- 33K tokens
- Input price
- 27.5 credits / 1M tokens ($2.50 raw)
- Output price
- 55.0 credits / 1M tokens ($5.00 raw)
- Vision (image input)
- No
- Tool calling
- No
- Extended reasoning
- No
Credits are what Zeplik bills: 1 credit = $0.10, computed from the raw provider rate with a 1.10x margin. Raw prices shown per 1M tokens.
Try Magnum v4 72B now
What Magnum v4 72B is best for
- Hard problems where answer quality matters more than cost: strategy, analysis, difficult writing
Frequently asked questions
- What is Magnum v4 72B?
- Magnum v4 72B is an AI model by Anthracite. This is a series of models designed to replicate the prose quality of the Claude 3 models, specifically Sonnet(https://openrouter.ai/anthropic/claude-3.5-sonnet) and Opus(https://openrouter.ai/anthropic/claude-3-opus). The model is fine-tuned on top of [Qwen2.5 72B](https://openrouter.ai/qwen/qwen-2.5-72b-instruct).
- How much does Magnum v4 72B cost on Zeplik?
- Input tokens cost 27.5 credits per million and output tokens 55.0 credits per million (1 credit = $0.10; the raw provider rates are $2.50 and $5.00 per million). New accounts start with free credits.
- How long can a conversation with Magnum v4 72B be?
- Magnum v4 72B has a 33K-token context window (32,768 tokens), which covers the conversation plus any documents you attach.
- Does Magnum v4 72B support images and tools?
- Magnum v4 72B is text-only and does not support tool calling.
Related models
GPT-6 Astra is OpenAI's flagship model for demanding end-to-end work. It is suited for advanced analysis, software engineering, deep research, scientific work, and document creation, with particular strengths in long-horizon...
GPT-6 Astra Pro is the same underlying model as [GPT-6 Astra](https://openrouter.ai/openai/gpt-6-astra), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode
Ling 3.0 Flash Sante is a health and medicine-focused mixture-of-experts model from InclusionAI, built on Ling 3.0 Flash with 5.1B active parameters out of 124B total. It is designed for...
Gemini 3.8 Flash is Google's most intelligent Flash model with significant gains from 3.7 Flash across software engineering, agentic tasks, and multi-step reasoning.
Gemini 3.8 Flash is Google's most intelligent Flash model with significant gains from 3.7 Flash across software engineering, agentic tasks, and multi-step reasoning.
Claude Fable 5.1 improves on Claude Fable 5 across the board, with the biggest gains in agentic coding, long-running agentic workflows, and knowledge work: long code refactors, front-end and visual...
More you can do with Magnum v4 72B
- 465 AI skillsReady-to-run expert methods for writing, data, code and more, each running on Magnum v4 72B or any model you pick.
- 997 integrationsConnect Gmail, Slack, GitHub, Notion and hundreds more so the assistant can work in your apps.
- Simple pricingPlus, Pro and Max plans with a monthly credit allowance that works across every model.