Gemma 3 4B
Open weightVisionGoogle · Released March 2025 · Knowledge cutoff August 2024
Gemma 3 introduces multimodality, supporting vision-language input and text outputs. It handles context windows up to 128k tokens, understands over 140 languages, and offers improved math, reasoning, and chat capabilities,...
It is part of Google's Gemini and Gemma catalog: Pro tiers lead on capability, Flash tiers on speed and price, and Gemma models are open weight.
Facts and pricing
- Provider
- Context window
- 131K tokens
- Input price
- 0.55 credits / 1M tokens ($0.050 raw)
- Output price
- 1.1 credits / 1M tokens ($0.10 raw)
- Vision (image input)
- Yes
- Tool calling
- No
- Extended reasoning
- No
Credits are what Zeplik bills: 1 credit = $0.10, computed from the raw provider rate with a 1.10x margin. Raw prices shown per 1M tokens.
Try Gemma 3 4B now
What Gemma 3 4B is best for
- Everyday chat and drafting on an open-weight model with transparent lineage
- Working with images: screenshots, charts, photos and scanned documents alongside text
Example prompts
Prompts that suit a open weight model like Gemma 3 4B:
- Draft a project update from these rough notes
- Explain how HTTPS works to a curious teenager
- Turn this list of features into a changelog entry
- Brainstorm objections to this proposal and how to answer them
Google family
| Model | Context | Input cr/M | Output cr/M | Released |
|---|---|---|---|---|
| Gemini 3.8 Flash (batch) | 1.0M | 4.1 | 20.6 | September 2026 |
| Gemini 3.8 Flash | 1.0M | 8.3 | 41.3 | September 2026 |
| Gemini 3.7 Flash | 1.0M | 8.3 | 41.3 | August 2026 |
| Gemini 3.7 Flash (batch) | 1.0M | 4.1 | 20.6 | August 2026 |
| Gemini 3.6 Flash | 1.0M | 8.3 | 41.3 | July 2026 |
| Gemini 3.6 Flash (batch) | 1.0M | 4.1 | 20.6 | July 2026 |
| Gemini 3.5 Flash Lite (batch) | 1.0M | 1.7 | 13.8 | July 2026 |
| Gemini 3.5 Flash Lite | 1.0M | 3.3 | 27.5 | July 2026 |
| Nano Banana 2 Lite (Gemini 3.1 Flash Lite Image) | 66K | 2.8 | 16.5 | June 2026 |
| Nano Banana 2 (Gemini 3.1 Flash Image) | 131K | 5.5 | 33.0 | June 2026 |
Frequently asked questions
- What is Gemma 3 4B?
- Gemma 3 4B is an AI model by Google. Gemma 3 introduces multimodality, supporting vision-language input and text outputs. It handles context windows up to 128k tokens, understands over 140 languages, and offers improved math, reasoning, and chat capabilities,...
- How much does Gemma 3 4B cost on Zeplik?
- Input tokens cost 0.55 credits per million and output tokens 1.1 credits per million (1 credit = $0.10; the raw provider rates are $0.050 and $0.10 per million). New accounts start with free credits.
- How long can a conversation with Gemma 3 4B be?
- Gemma 3 4B has a 131K-token context window (131,072 tokens), which covers the conversation plus any documents you attach.
- Does Gemma 3 4B support images and tools?
- Gemma 3 4B accepts image input and does not support tool calling.
Related models
Gemini 3.8 Flash is Google's most intelligent Flash model with significant gains from 3.7 Flash across software engineering, agentic tasks, and multi-step reasoning.
Gemini 3.8 Flash is Google's most intelligent Flash model with significant gains from 3.7 Flash across software engineering, agentic tasks, and multi-step reasoning.
Gemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning. It is designed for tasks that require responsive performance and reliable multi-step...
Gemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning. It is designed for tasks that require responsive performance and reliable multi-step...
Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and...
Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and...
More you can do with Gemma 3 4B
- 465 AI skillsReady-to-run expert methods for writing, data, code and more, each running on Gemma 3 4B or any model you pick.
- 997 integrationsConnect Gmail, Slack, GitHub, Notion and hundreds more so the assistant can work in your apps.
- Simple pricingPlus, Pro and Max plans with a monthly credit allowance that works across every model.