Gemma 4 26B A4B (free)
Open weightVisionToolsGoogle · Released April 2026
Gemma 4 26B A4B IT is an instruction-tuned Mixture-of-Experts (MoE) model from Google DeepMind. Despite 25.2B total parameters, only 3.8B activate per token during inference — delivering near-31B quality at...
It is part of Google's Gemini and Gemma catalog: Pro tiers lead on capability, Flash tiers on speed and price, and Gemma models are open weight.
Facts and pricing
- Provider
- Context window
- 262K tokens
- Input price
- 0 credits / 1M tokens ($0 raw)
- Output price
- 0 credits / 1M tokens ($0 raw)
- Vision (image input)
- Yes
- Tool calling
- Yes
- Extended reasoning
- No
Credits are what Zeplik bills: 1 credit = $0.10, computed from the raw provider rate with a 1.10x margin. Raw prices shown per 1M tokens.
Try Gemma 4 26B A4B now
Gemma 4 26B A4B (free) is a serving variant of Gemma 4 26B A4B, so Zeplik offers the Gemma 4 26B A4B row and routes your prompt there.
What Gemma 4 26B A4B (free) is best for
- Everyday chat and drafting on an open-weight model with transparent lineage
- Long documents: the 262K-token window holds lengthy reports, contracts or papers whole
- Working with images: screenshots, charts, photos and scanned documents alongside text
- Tool use and agents: reliably calls functions, so it can search, run skills and drive workflows
Example prompts
Prompts that suit a open weight model like Gemma 4 26B A4B (free):
- Draft a project update from these rough notes
- Explain how HTTPS works to a curious teenager
- Turn this list of features into a changelog entry
- Brainstorm objections to this proposal and how to answer them
Google family
| Model | Context | Input cr/M | Output cr/M | Released |
|---|---|---|---|---|
| Gemini 3.8 Flash (batch) | 1.0M | 4.1 | 20.6 | September 2026 |
| Gemini 3.8 Flash | 1.0M | 8.3 | 41.3 | September 2026 |
| Gemini 3.7 Flash | 1.0M | 8.3 | 41.3 | August 2026 |
| Gemini 3.7 Flash (batch) | 1.0M | 4.1 | 20.6 | August 2026 |
| Gemini 3.6 Flash | 1.0M | 8.3 | 41.3 | July 2026 |
| Gemini 3.6 Flash (batch) | 1.0M | 4.1 | 20.6 | July 2026 |
| Gemini 3.5 Flash Lite (batch) | 1.0M | 1.7 | 13.8 | July 2026 |
| Gemini 3.5 Flash Lite | 1.0M | 3.3 | 27.5 | July 2026 |
| Nano Banana 2 Lite (Gemini 3.1 Flash Lite Image) | 66K | 2.8 | 16.5 | June 2026 |
| Nano Banana 2 (Gemini 3.1 Flash Image) | 131K | 5.5 | 33.0 | June 2026 |
Frequently asked questions
- What is Gemma 4 26B A4B (free)?
- Gemma 4 26B A4B (free) is an AI model by Google. Gemma 4 26B A4B IT is an instruction-tuned Mixture-of-Experts (MoE) model from Google DeepMind. Despite 25.2B total parameters, only 3.8B activate per token during inference — delivering near-31B quality at...
- How much does Gemma 4 26B A4B (free) cost on Zeplik?
- Input tokens cost 0 credits per million and output tokens 0 credits per million (1 credit = $0.10; the raw provider rates are $0 and $0 per million). New accounts start with free credits.
- How long can a conversation with Gemma 4 26B A4B (free) be?
- Gemma 4 26B A4B (free) has a 262K-token context window (262,144 tokens), which covers the conversation plus any documents you attach.
- Does Gemma 4 26B A4B (free) support images and tools?
- Gemma 4 26B A4B (free) accepts image input and supports tool calling.
Related models
Gemini 3.8 Flash is Google's most intelligent Flash model with significant gains from 3.7 Flash across software engineering, agentic tasks, and multi-step reasoning.
Gemini 3.8 Flash is Google's most intelligent Flash model with significant gains from 3.7 Flash across software engineering, agentic tasks, and multi-step reasoning.
Gemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning. It is designed for tasks that require responsive performance and reliable multi-step...
Gemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning. It is designed for tasks that require responsive performance and reliable multi-step...
Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and...
Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and...
More you can do with Gemma 4 26B A4B (free)
- 465 AI skillsReady-to-run expert methods for writing, data, code and more, each running on Gemma 4 26B A4B (free) or any model you pick.
- 997 integrationsConnect Gmail, Slack, GitHub, Notion and hundreds more so the assistant can work in your apps.
- Simple pricingPlus, Pro and Max plans with a monthly credit allowance that works across every model.