Granite 4.2 8B
Open weightToolsIBM · Released August 2026
Granite 4.2 8B is a dense reasoning model from IBM. It is suited for mathematics, code generation, multilingual dialogue, and agentic workflows that need multi-step reasoning. It supports full, low-effort,...
Facts and pricing
- Provider
- IBM
- Context window
- 131K tokens
- Input price
- 1.1 credits / 1M tokens ($0.10 raw)
- Output price
- 1.7 credits / 1M tokens ($0.15 raw)
- Vision (image input)
- No
- Tool calling
- Yes
- Extended reasoning
- No
Credits are what Zeplik bills: 1 credit = $0.10, computed from the raw provider rate with a 1.10x margin. Raw prices shown per 1M tokens.
Try Granite 4.2 8B now
What Granite 4.2 8B is best for
- Everyday chat and drafting on an open-weight model with transparent lineage
- Tool use and agents: reliably calls functions, so it can search, run skills and drive workflows
IBM family
| Model | Context | Input cr/M | Output cr/M | Released |
|---|---|---|---|---|
| Granite 4.2 8B | 131K | 1.1 | 1.7 | August 2026 |
| Granite 4.0 Micro | 131K | 0.19 | 1.2 | October 2025 |
Frequently asked questions
- What is Granite 4.2 8B?
- Granite 4.2 8B is an AI model by IBM. Granite 4.2 8B is a dense reasoning model from IBM. It is suited for mathematics, code generation, multilingual dialogue, and agentic workflows that need multi-step reasoning. It supports full, low-effort,...
- How much does Granite 4.2 8B cost on Zeplik?
- Input tokens cost 1.1 credits per million and output tokens 1.7 credits per million (1 credit = $0.10; the raw provider rates are $0.10 and $0.15 per million). New accounts start with free credits.
- How long can a conversation with Granite 4.2 8B be?
- Granite 4.2 8B has a 131K-token context window (131,072 tokens), which covers the conversation plus any documents you attach.
- Does Granite 4.2 8B support images and tools?
- Granite 4.2 8B is text-only and supports tool calling.
Related models
Granite-4.0-H-Micro is a 3B parameter from the Granite 4 family of models. These models are the latest in a series of models released by IBM. They are fine-tuned for long...
GPT-6 Astra is OpenAI's flagship model for demanding end-to-end work. It is suited for advanced analysis, software engineering, deep research, scientific work, and document creation, with particular strengths in long-horizon...
GPT-6 Astra Pro is the same underlying model as [GPT-6 Astra](https://openrouter.ai/openai/gpt-6-astra), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode
Ling 3.0 Flash Sante is a health and medicine-focused mixture-of-experts model from InclusionAI, built on Ling 3.0 Flash with 5.1B active parameters out of 124B total. It is designed for...
Gemini 3.8 Flash is Google's most intelligent Flash model with significant gains from 3.7 Flash across software engineering, agentic tasks, and multi-step reasoning.
Gemini 3.8 Flash is Google's most intelligent Flash model with significant gains from 3.7 Flash across software engineering, agentic tasks, and multi-step reasoning.
More you can do with Granite 4.2 8B
- 465 AI skillsReady-to-run expert methods for writing, data, code and more, each running on Granite 4.2 8B or any model you pick.
- 997 integrationsConnect Gmail, Slack, GitHub, Notion and hundreds more so the assistant can work in your apps.
- Simple pricingPlus, Pro and Max plans with a monthly credit allowance that works across every model.