google

Google: Gemini 3.1 Flash Lite

Gemini 3.1 Flash Lite is a multimodal model from Google that accepts text, images, video, audio, and files as input. Its context window reaches 1,048,576 tokens, making it suitable for tasks that require ingesting long documents or extended conversation histories. The model supports tool use and reasoning, which enables agentic workflows and multi-step problem solving. Maximum output is capped at 65,536 tokens per response. At $0.25 per million input tokens and $1.50 per million output tokens, it sits at the affordable end of the market, which makes it worth considering for high-volume or cost-sensitive applications. The tradeoff is that there is currently no independent benchmark coverage to validate its quality claims, so teams evaluating Gemini 3.1 Flash Lite are working without third-party performance data. Buyers who prioritize proven, measurable quality should treat it as unproven for now and run their own evals before committing.

Quality Score
100/100
price + capability + benchmarks
Input Price
$0.25
per 1M tokens
Output Price
$1.50
per 1M tokens
Context Window
1,048,576
tokens
Model ID
google/gemini-3.1-flash-lite
Vendor
google
Released
May 2026
Tokenizer
Gemini
Input Modalities
text, image, video, file, audio
Output Modalities
text
Max Output
65,536 tokens
Tool Calling
✓ supported
Structured Output
✓ supported
Reasoning Mode
✓ supported
Vision
✓ accepts images
Audio
✓ accepts audio
Moderated
no

What it costs in practice

Computed from the current $0.25/M input and $1.50/M output rates. Run your own numbers →

JobTokensCost
Summarize a 50-page report 30k in / 1.5k out under $0.01
Classify 1,000 customer emails 500k in / 50k out $0.20
A month of a busy support chatbot 5M in / 2M out $4.25

Category rankings

Where Google: Gemini 3.1 Flash Lite places across the 4 categories it ranks in. How we rank →

#CategoryScore
#8 TranscriptionVoice · of 25 ranked 123
#11 Video Auto-TaggingVideo · of 25 ranked 123
#12 TTS ReplacementVoice · of 25 ranked 115
#23 Audio SummarizationVoice · of 25 ranked 139

Similar models

Quick answers

How much does Google: Gemini 3.1 Flash Lite cost?
$0.25 per million input tokens and $1.50 per million output tokens.
What is Google: Gemini 3.1 Flash Lite's context window?
1,048,576 tokens, roughly 1,572 pages of text in a single request.
Does Google: Gemini 3.1 Flash Lite support tool calling?
It supports tool calling, structured output, a reasoning mode.
Can Google: Gemini 3.1 Flash Lite process images?
Yes, it accepts image input.