Google: Gemini 3.1 Flash Lite
Gemini 3.1 Flash Lite is a multimodal model from Google that accepts text, images, video, audio, and files as input. Its context window reaches 1,048,576 tokens, making it suitable for tasks that require ingesting long documents or extended conversation histories. The model supports tool use and reasoning, which enables agentic workflows and multi-step problem solving. Maximum output is capped at 65,536 tokens per response. At $0.25 per million input tokens and $1.50 per million output tokens, it sits at the affordable end of the market, which makes it worth considering for high-volume or cost-sensitive applications. The tradeoff is that there is currently no independent benchmark coverage to validate its quality claims, so teams evaluating Gemini 3.1 Flash Lite are working without third-party performance data. Buyers who prioritize proven, measurable quality should treat it as unproven for now and run their own evals before committing.
- Model ID
- google/gemini-3.1-flash-lite
- Vendor
- Released
- May 2026
- Tokenizer
- Gemini
- Input Modalities
- text, image, video, file, audio
- Output Modalities
- text
- Max Output
- 65,536 tokens
- Tool Calling
- ✓ supported
- Structured Output
- ✓ supported
- Reasoning Mode
- ✓ supported
- Vision
- ✓ accepts images
- Audio
- ✓ accepts audio
- Moderated
- no
What it costs in practice
Computed from the current $0.25/M input and $1.50/M output rates. Run your own numbers →
| Job | Tokens | Cost |
|---|---|---|
| Summarize a 50-page report | 30k in / 1.5k out | under $0.01 |
| Classify 1,000 customer emails | 500k in / 50k out | $0.20 |
| A month of a busy support chatbot | 5M in / 2M out | $4.25 |
Category rankings
Where Google: Gemini 3.1 Flash Lite places across the 4 categories it ranks in. How we rank →
| # | Category | Score |
|---|---|---|
| #8 | TranscriptionVoice · of 25 ranked | 123 |
| #11 | Video Auto-TaggingVideo · of 25 ranked | 123 |
| #12 | TTS ReplacementVoice · of 25 ranked | 115 |
| #23 | Audio SummarizationVoice · of 25 ranked | 139 |
Similar models
Google: Gemini 3.6 Flash
Google: Gemini 3.6 Flash (batch)
Google: Gemini 3.5 Flash Lite
Google: Gemini 3.5 Flash Lite (batch)
Google: Gemini 3.5 Flash
Google: Gemini 3.5 Flash (batch)
Quick answers
- How much does Google: Gemini 3.1 Flash Lite cost?
- $0.25 per million input tokens and $1.50 per million output tokens.
- What is Google: Gemini 3.1 Flash Lite's context window?
- 1,048,576 tokens, roughly 1,572 pages of text in a single request.
- Does Google: Gemini 3.1 Flash Lite support tool calling?
- It supports tool calling, structured output, a reasoning mode.
- Can Google: Gemini 3.1 Flash Lite process images?
- Yes, it accepts image input.