google

Google: Gemini 3.1 Flash Lite (batch)

Google’s Gemini 3.1 Flash Lite is a powerful AI model that supports text, images, videos, files, and audio inputs; it can process context up to 1MB in length and facilitate reasoning processes. This model also allows for the integration of tools, enhancing its utility across various applications. However, structured output is not supported, which limits certain use cases. Given its price at $0.125 per million input tokens and $0.75 per million output tokens, Gemini 3.1 Flash Lite may be more suitable for users who have a specific budget and need to handle diverse modalities but do not require structured outputs. Until independent benchmarks are available, its performance remains unverified in comparative evaluations.

Quality Score
100/100
price + capability + benchmarks
Input Price
$0.12
per 1M tokens
Output Price
$0.75
per 1M tokens
Context Window
1,048,576
tokens
Model ID
google/gemini-3.1-flash-lite:batch
Vendor
google
Released
May 2026
Tokenizer
Gemini
Input Modalities
text, image, video, file, audio
Output Modalities
text
Max Output
65,536 tokens
Tool Calling
✓ supported
Structured Output
✓ supported
Reasoning Mode
✓ supported
Vision
✓ accepts images
Audio
✓ accepts audio
Moderated
no

What it costs in practice

Computed from the current $0.12/M input and $0.75/M output rates. Run your own numbers →

JobTokensCost
Summarize a 50-page report 30k in / 1.5k out under $0.01
Classify 1,000 customer emails 500k in / 50k out $0.10
A month of a busy support chatbot 5M in / 2M out $2.12

Category rankings

Where Google: Gemini 3.1 Flash Lite (batch) places across the 5 categories it ranks in. How we rank →

#CategoryScore
#9 TranscriptionVoice · of 25 ranked 123
#12 Video Auto-TaggingVideo · of 25 ranked 123
#13 TTS ReplacementVoice · of 25 ranked 115
#24 Audio SummarizationVoice · of 25 ranked 139
#25 Real-Time ChatLatency · of 25 ranked 117

Similar models

Quick answers

How much does Google: Gemini 3.1 Flash Lite (batch) cost?
$0.12 per million input tokens and $0.75 per million output tokens.
What is Google: Gemini 3.1 Flash Lite (batch)'s context window?
1,048,576 tokens, roughly 1,572 pages of text in a single request.
Does Google: Gemini 3.1 Flash Lite (batch) support tool calling?
It supports tool calling, structured output, a reasoning mode.
Can Google: Gemini 3.1 Flash Lite (batch) process images?
Yes, it accepts image input.