Google: Gemini 2.5 Flash Lite (batch)
Gemini 2.5 Flash Lite by Google is designed for extensive text, image, file, audio, and video processing with a context length of up to 1,048,576 tokens. It supports reasoning capabilities and tool integration, making it versatile for complex tasks. This model should be considered for projects requiring comprehensive multimodal handling, given its broad input modalities and reasoning support. Its blended benchmark score of 9.6 places it among the top performers; however, its pricing at $0.05 per million input tokens and $0.2 per million output tokens makes it somewhat pricey compared to other models. Despite its strong performance, the lack of structured output and open weights may limit certain use cases, but for those needing a robust toolset and high context handling, Gemini 2.5 Flash Lite remains a viable choice.
Benchmark results
Independent, published benchmarks. Blended score 9.6 across 1 benchmark, last refreshed 2026-08-13. How scoring works →
| Benchmark | Measures | Score |
|---|---|---|
| AI Index | broad capability composite | 11.0 |
- Model ID
- google/gemini-2.5-flash-lite:batch
- Vendor
- Released
- July 2025
- Tokenizer
- Gemini
- Input Modalities
- text, image, file, audio, video
- Output Modalities
- text
- Max Output
- 65,535 tokens
- Tool Calling
- ✓ supported
- Structured Output
- ✓ supported
- Reasoning Mode
- ✓ supported
- Vision
- ✓ accepts images
- Audio
- ✓ accepts audio
- Moderated
- no
What it costs in practice
Computed from the current $0.05/M input and $0.20/M output rates. Run your own numbers →
| Job | Tokens | Cost |
|---|---|---|
| Summarize a 50-page report | 30k in / 1.5k out | under $0.01 |
| Classify 1,000 customer emails | 500k in / 50k out | $0.04 |
| A month of a busy support chatbot | 5M in / 2M out | $0.65 |
Category rankings
Where Google: Gemini 2.5 Flash Lite (batch) places across the 8 categories it ranks in. How we rank →
| # | Category | Score |
|---|---|---|
| #8 | Social Media PostsWriting · of 25 ranked | 120 |
| #8 | Voice Assistant BackendVoice · of 25 ranked | 124 |
| #8 | Cheap Bulk InferenceCost · of 25 ranked | 137 |
| #13 | Self-Hosted / LocalCost · of 25 ranked | 117 |
| #15 | Real-Time ChatLatency · of 25 ranked | 118 |
| #16 | TranscriptionVoice · of 25 ranked | 123 |
| #22 | Audio SummarizationVoice · of 25 ranked | 140 |
| #25 | TTS ReplacementVoice · of 25 ranked | 115 |
Similar models
Google: Gemini 3.6 Flash
Google: Gemini 3.6 Flash (batch)
Google: Gemini 3.5 Flash Lite
Google: Gemini 3.5 Flash Lite (batch)
Google: Gemini 3.5 Flash
Google: Gemini 3.5 Flash (batch)
Quick answers
- How much does Google: Gemini 2.5 Flash Lite (batch) cost?
- $0.05 per million input tokens and $0.20 per million output tokens.
- What is Google: Gemini 2.5 Flash Lite (batch)'s context window?
- 1,048,576 tokens, roughly 1,572 pages of text in a single request.
- Does Google: Gemini 2.5 Flash Lite (batch) support tool calling?
- It supports tool calling, structured output, a reasoning mode.
- Can Google: Gemini 2.5 Flash Lite (batch) process images?
- Yes, it accepts image input.