Gemini 3.1 Flash Lite
Fast and lightVisionToolsGoogle · Released May 2026
Gemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads. It supports text, image, video, audio, and PDF inputs, and is designed for lightweight agentic...
It is part of Google's Gemini and Gemma catalog: Pro tiers lead on capability, Flash tiers on speed and price, and Gemma models are open weight.
Facts and pricing
- Provider
- Context window
- 1.0M tokens
- Input price
- 2.8 credits / 1M tokens ($0.25 raw)
- Output price
- 16.5 credits / 1M tokens ($1.50 raw)
- Vision (image input)
- Yes
- Tool calling
- Yes
- Extended reasoning
- No
Credits are what Zeplik bills: 1 credit = $0.10, computed from the raw provider rate with a 1.10x margin. Raw prices shown per 1M tokens.
Try Gemini 3.1 Flash Lite now
What Gemini 3.1 Flash Lite is best for
- High-volume everyday tasks where speed and cost matter: summaries, drafts, quick questions
- Very long documents and codebases: a 1.0M-token window fits entire books or repositories in one conversation
- Working with images: screenshots, charts, photos and scanned documents alongside text
- Tool use and agents: reliably calls functions, so it can search, run skills and drive workflows
Example prompts
Prompts that suit a fast and light model like Gemini 3.1 Flash Lite:
- Summarize this article in five bullet points
- Rewrite this email to be shorter and friendlier
- Give me ten name ideas for a hiking newsletter
- Translate this paragraph to Spanish and keep the tone
Google family
| Model | Context | Input cr/M | Output cr/M | Released |
|---|---|---|---|---|
| Gemini 3.8 Flash (batch) | 1.0M | 4.1 | 20.6 | September 2026 |
| Gemini 3.8 Flash | 1.0M | 8.3 | 41.3 | September 2026 |
| Gemini 3.7 Flash | 1.0M | 8.3 | 41.3 | August 2026 |
| Gemini 3.7 Flash (batch) | 1.0M | 4.1 | 20.6 | August 2026 |
| Gemini 3.6 Flash | 1.0M | 8.3 | 41.3 | July 2026 |
| Gemini 3.6 Flash (batch) | 1.0M | 4.1 | 20.6 | July 2026 |
| Gemini 3.5 Flash Lite (batch) | 1.0M | 1.7 | 13.8 | July 2026 |
| Gemini 3.5 Flash Lite | 1.0M | 3.3 | 27.5 | July 2026 |
| Nano Banana 2 Lite (Gemini 3.1 Flash Lite Image) | 66K | 2.8 | 16.5 | June 2026 |
| Nano Banana 2 (Gemini 3.1 Flash Image) | 131K | 5.5 | 33.0 | June 2026 |
Compare Gemini 3.1 Flash Lite
Frequently asked questions
- What is Gemini 3.1 Flash Lite?
- Gemini 3.1 Flash Lite is an AI model by Google. Gemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads. It supports text, image, video, audio, and PDF inputs, and is designed for lightweight agentic...
- How much does Gemini 3.1 Flash Lite cost on Zeplik?
- Input tokens cost 2.8 credits per million and output tokens 16.5 credits per million (1 credit = $0.10; the raw provider rates are $0.25 and $1.50 per million). New accounts start with free credits.
- How long can a conversation with Gemini 3.1 Flash Lite be?
- Gemini 3.1 Flash Lite has a 1.0M-token context window (1,048,576 tokens), which covers the conversation plus any documents you attach.
- Does Gemini 3.1 Flash Lite support images and tools?
- Gemini 3.1 Flash Lite accepts image input and supports tool calling.
Related models
Gemini 3.8 Flash is Google's most intelligent Flash model with significant gains from 3.7 Flash across software engineering, agentic tasks, and multi-step reasoning.
Gemini 3.8 Flash is Google's most intelligent Flash model with significant gains from 3.7 Flash across software engineering, agentic tasks, and multi-step reasoning.
Gemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning. It is designed for tasks that require responsive performance and reliable multi-step...
Gemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning. It is designed for tasks that require responsive performance and reliable multi-step...
Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and...
Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and...
More you can do with Gemini 3.1 Flash Lite
- 465 AI skillsReady-to-run expert methods for writing, data, code and more, each running on Gemini 3.1 Flash Lite or any model you pick.
- 997 integrationsConnect Gmail, Slack, GitHub, Notion and hundreds more so the assistant can work in your apps.
- Simple pricingPlus, Pro and Max plans with a monthly credit allowance that works across every model.