Gemma 4 31B (free)
Open weightVisionToolsGoogle · Released April 2026
Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function...
It is part of Google's Gemini and Gemma catalog: Pro tiers lead on capability, Flash tiers on speed and price, and Gemma models are open weight.
Facts and pricing
- Provider
- Context window
- 262K tokens
- Input price
- 0 credits / 1M tokens ($0 raw)
- Output price
- 0 credits / 1M tokens ($0 raw)
- Vision (image input)
- Yes
- Tool calling
- Yes
- Extended reasoning
- No
Credits are what Zeplik bills: 1 credit = $0.10, computed from the raw provider rate with a 1.10x margin. Raw prices shown per 1M tokens.
Try Gemma 4 31B now
Gemma 4 31B (free) is a serving variant of Gemma 4 31B, so Zeplik offers the Gemma 4 31B row and routes your prompt there.
What Gemma 4 31B (free) is best for
- Everyday chat and drafting on an open-weight model with transparent lineage
- Long documents: the 262K-token window holds lengthy reports, contracts or papers whole
- Working with images: screenshots, charts, photos and scanned documents alongside text
- Tool use and agents: reliably calls functions, so it can search, run skills and drive workflows
Example prompts
Prompts that suit a open weight model like Gemma 4 31B (free):
- Draft a project update from these rough notes
- Explain how HTTPS works to a curious teenager
- Turn this list of features into a changelog entry
- Brainstorm objections to this proposal and how to answer them
Google family
| Model | Context | Input cr/M | Output cr/M | Released |
|---|---|---|---|---|
| Gemini 3.8 Flash (batch) | 1.0M | 4.1 | 20.6 | September 2026 |
| Gemini 3.8 Flash | 1.0M | 8.3 | 41.3 | September 2026 |
| Gemini 3.7 Flash | 1.0M | 8.3 | 41.3 | August 2026 |
| Gemini 3.7 Flash (batch) | 1.0M | 4.1 | 20.6 | August 2026 |
| Gemini 3.6 Flash | 1.0M | 8.3 | 41.3 | July 2026 |
| Gemini 3.6 Flash (batch) | 1.0M | 4.1 | 20.6 | July 2026 |
| Gemini 3.5 Flash Lite (batch) | 1.0M | 1.7 | 13.8 | July 2026 |
| Gemini 3.5 Flash Lite | 1.0M | 3.3 | 27.5 | July 2026 |
| Nano Banana 2 Lite (Gemini 3.1 Flash Lite Image) | 66K | 2.8 | 16.5 | June 2026 |
| Nano Banana 2 (Gemini 3.1 Flash Image) | 131K | 5.5 | 33.0 | June 2026 |
Frequently asked questions
- What is Gemma 4 31B (free)?
- Gemma 4 31B (free) is an AI model by Google. Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function...
- How much does Gemma 4 31B (free) cost on Zeplik?
- Input tokens cost 0 credits per million and output tokens 0 credits per million (1 credit = $0.10; the raw provider rates are $0 and $0 per million). New accounts start with free credits.
- How long can a conversation with Gemma 4 31B (free) be?
- Gemma 4 31B (free) has a 262K-token context window (262,144 tokens), which covers the conversation plus any documents you attach.
- Does Gemma 4 31B (free) support images and tools?
- Gemma 4 31B (free) accepts image input and supports tool calling.
Related models
Gemini 3.8 Flash is Google's most intelligent Flash model with significant gains from 3.7 Flash across software engineering, agentic tasks, and multi-step reasoning.
Gemini 3.8 Flash is Google's most intelligent Flash model with significant gains from 3.7 Flash across software engineering, agentic tasks, and multi-step reasoning.
Gemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning. It is designed for tasks that require responsive performance and reliable multi-step...
Gemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning. It is designed for tasks that require responsive performance and reliable multi-step...
Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and...
Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and...
More you can do with Gemma 4 31B (free)
- 465 AI skillsReady-to-run expert methods for writing, data, code and more, each running on Gemma 4 31B (free) or any model you pick.
- 997 integrationsConnect Gmail, Slack, GitHub, Notion and hundreds more so the assistant can work in your apps.
- Simple pricingPlus, Pro and Max plans with a monthly credit allowance that works across every model.