Google: Gemma 3 12B
Gemma 3 12B from Google is designed for handling text and image inputs with a context length of up to 131072 tokens and supports tools integration but lacks reasoning capabilities or structured output options. It processes both modalities effectively, making it suitable for tasks that require interaction between textual and visual data. While Gemma 3 offers a solid benchmark score of 4.6 across three independent tests, its performance in agentic tasks is notably weak, scoring only 0.5 on the AI Index. Given its pricing at $0.05 per thousand input tokens and $0.15 per thousand output tokens, it may be more economical for users who frequently interact with images or require tool integration, though its limited reasoning abilities might restrict its usefulness in dynamic problem-solving scenarios.
Benchmark results
Independent, published benchmarks. Blended score 4.6 across 3 benchmarks, last refreshed 2026-08-13. How scoring works →
| Benchmark | Measures | Score |
|---|---|---|
| AI Index | broad capability composite | 9.1 |
| AI Index Coding | software engineering tasks | 9.5 |
| AI Index Agentic | multi-step tool-using tasks | 0.5 |
- Model ID
- google/gemma-3-12b-it
- Vendor
- Released
- March 2025
- Tokenizer
- Gemini
- Input Modalities
- text, image
- Output Modalities
- text
- Max Output
- 16,384 tokens
- Tool Calling
- ✓ supported
- Structured Output
- ✓ supported
- Reasoning Mode
- not supported
- Vision
- ✓ accepts images
- Audio
- no
- Moderated
- no
What it costs in practice
Computed from the current $0.05/M input and $0.15/M output rates. Run your own numbers →
| Job | Tokens | Cost |
|---|---|---|
| Summarize a 50-page report | 30k in / 1.5k out | under $0.01 |
| Classify 1,000 customer emails | 500k in / 50k out | $0.03 |
| A month of a busy support chatbot | 5M in / 2M out | $0.55 |
Similar models
Google: Nano Banana Pro (Gemini 3 Pro Image)
Google: Lyria 3 Pro Preview
Google: Lyria 3 Clip Preview
Google: Nano Banana 2 (Gemini 3.1 Flash Image)
Google: Gemma 3 27B
Google: Gemini 3.6 Flash
Quick answers
- How much does Google: Gemma 3 12B cost?
- $0.05 per million input tokens and $0.15 per million output tokens.
- What is Google: Gemma 3 12B's context window?
- 131,072 tokens, roughly 196 pages of text in a single request.
- Does Google: Gemma 3 12B support tool calling?
- It supports tool calling, structured output.
- Can Google: Gemma 3 12B process images?
- Yes, it accepts image input.