Skip to main content

Granite 4.0 Micro

Open weight

IBM · Released October 2025

Granite-4.0-H-Micro is a 3B parameter from the Granite 4 family of models. These models are the latest in a series of models released by IBM. They are fine-tuned for long...

Facts and pricing

Provider
IBM
Context window
131K tokens
Input price
0.19 credits / 1M tokens ($0.017 raw)
Output price
1.2 credits / 1M tokens ($0.11 raw)
Vision (image input)
No
Tool calling
No
Extended reasoning
No

Credits are what Zeplik bills: 1 credit = $0.10, computed from the raw provider rate with a 1.10x margin. Raw prices shown per 1M tokens.

Try Granite 4.0 Micro now

Ask Granite 4.0 Micro anything. Your prompt opens in the Zeplik app with this model selected.

Granite 4.0 Micro

Or open Zeplik with Granite 4.0 Micro already selected

What Granite 4.0 Micro is best for

IBM family

ModelContextInput cr/MOutput cr/MReleased
Granite 4.2 8B131K1.11.7August 2026
Granite 4.0 Micro131K0.191.2October 2025

Frequently asked questions

What is Granite 4.0 Micro?
Granite 4.0 Micro is an AI model by IBM. Granite-4.0-H-Micro is a 3B parameter from the Granite 4 family of models. These models are the latest in a series of models released by IBM. They are fine-tuned for long...
How much does Granite 4.0 Micro cost on Zeplik?
Input tokens cost 0.19 credits per million and output tokens 1.2 credits per million (1 credit = $0.10; the raw provider rates are $0.017 and $0.11 per million). New accounts start with free credits.
How long can a conversation with Granite 4.0 Micro be?
Granite 4.0 Micro has a 131K-token context window (131,000 tokens), which covers the conversation plus any documents you attach.
Does Granite 4.0 Micro support images and tools?
Granite 4.0 Micro is text-only and does not support tool calling.

Related models

Granite 4.2 8BOpen weight

Granite 4.2 8B is a dense reasoning model from IBM. It is suited for mathematics, code generation, multilingual dialogue, and agentic workflows that need multi-step reasoning. It supports full, low-effort,...

131K ctx1.7 cr/M outAugust 2026
GPT-6 AstraFrontier

GPT-6 Astra is OpenAI's flagship model for demanding end-to-end work. It is suited for advanced analysis, software engineering, deep research, scientific work, and document creation, with particular strengths in long-horizon...

1.1M ctx550 cr/M outSeptember 2026
GPT-6 Astra ProFrontier

GPT-6 Astra Pro is the same underlying model as [GPT-6 Astra](https://openrouter.ai/openai/gpt-6-astra), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode

1.1M ctx550 cr/M outSeptember 2026
Ling 3.0 Flash Sante (free)Fast and light

Ling 3.0 Flash Sante is a health and medicine-focused mixture-of-experts model from InclusionAI, built on Ling 3.0 Flash with 5.1B active parameters out of 124B total. It is designed for...

262K ctx0 cr/M outSeptember 2026
Gemini 3.8 Flash (batch)Fast and light

Gemini 3.8 Flash is Google's most intelligent Flash model with significant gains from 3.7 Flash across software engineering, agentic tasks, and multi-step reasoning.

1.0M ctx20.6 cr/M outSeptember 2026
Gemini 3.8 FlashFast and light

Gemini 3.8 Flash is Google's most intelligent Flash model with significant gains from 3.7 Flash across software engineering, agentic tasks, and multi-step reasoning.

1.0M ctx41.3 cr/M outSeptember 2026

More you can do with Granite 4.0 Micro

Granite 4.0 Micro - IBM AI Model | Zeplik Chat