Mercury 2
General purposeToolsInception · Released March 2026
Mercury 2 is an extremely fast reasoning LLM, and the first reasoning diffusion LLM (dLLM).
It is part of Inception's Mercury line of diffusion language models, which generate tokens in parallel rather than sequentially and are built for speed.
Facts and pricing
- Provider
- Inception
- Context window
- 128K tokens
- Input price
- 3.0 credits / 1M tokens ($0.25 raw)
- Output price
- 9.0 credits / 1M tokens ($0.75 raw)
- Vision (image input)
- No
- Tool calling
- Yes
- Extended reasoning
- No
Credits are what Zeplik bills: 1 credit = $0.10, computed from the raw provider rate with a 1.20x margin. Raw prices shown per 1M tokens.
Try Mercury 2 now
What Mercury 2 is best for
- Balanced everyday work: writing, questions, brainstorming and summarization
- Tool use and agents: reliably calls functions, so it can search, run skills and drive workflows
Example prompts
Prompts that suit a general purpose model like Mercury 2:
- Help me outline a talk about what I learned this year
- Compare these two paragraphs and tell me which is clearer and why
- Draft a polite reply declining this invitation
- Explain this concept with a concrete everyday analogy
Inception family
| Model | Context | Input cr/M | Output cr/M | Released |
|---|---|---|---|---|
| Mercury 2.5 | 260K | 0.48 | 1.8 | September 2026 |
| Mercury 2 | 128K | 3.0 | 9.0 | March 2026 |
Frequently asked questions
- What is Mercury 2?
- Mercury 2 is an AI model by Inception. Mercury 2 is an extremely fast reasoning LLM, and the first reasoning diffusion LLM (dLLM).
- How much does Mercury 2 cost on Zeplik?
- Input tokens cost 3.0 credits per million and output tokens 9.0 credits per million (1 credit = $0.10; the raw provider rates are $0.25 and $0.75 per million). New accounts start with free credits.
- How long can a conversation with Mercury 2 be?
- Mercury 2 has a 128K-token context window (128,000 tokens), which covers the conversation plus any documents you attach.
- Does Mercury 2 support images and tools?
- Mercury 2 is text-only and supports tool calling.
Related models
Mercury 2.5 is the fastest reasoning LLM, and the latest diffusion LLM (dLLM) from Inception.
Claude Haiku 5.5 is Anthropic's small, fast model for high-volume, cost-sensitive work such as summarization, subagents, and browser use.
Nano Banana 2.1 (Gemini Nano Banana 2.1) is Google's image generation and editing model on the Flash tier, succeeding Nano Banana 2 and Nano Banana Pro.
Mistral Large 4 is a frontier multimodal (text and image input) model from Mistral AI built for reasoning, coding, and agentic workloads.
Ling 3.1 Flash is a hybrid reasoning mixture-of-experts model from inclusionAI, with 25B active parameters out of 560B total.
Apodex 1.1 Mini is a reasoning-first model from Apodex, built for complex, long-horizon research and forecasting tasks.
More you can do with Mercury 2
- 467 AI skillsReady-to-run expert methods for writing, data, code and more, each running on Mercury 2 or any model you pick.
- 997 integrationsConnect Gmail, Slack, GitHub, Notion and hundreds more so the assistant can work in your apps.
- Simple pricingPlus, Pro and Max plans with a monthly credit allowance that works across every model.