Qwen: Qwen3 VL 30B A3B Instruct
Qwen3 VL 30B A3B Instruct is capable of processing both text and images, with a context length of 262,144 tokens and support for external tools. It does not offer reasoning or structured output capabilities; however, it excels in handling multi-modal inputs effectively. This model is suitable for tasks that require understanding and processing of diverse data types, such as image captioning or text-image alignment. Given its blended benchmark score of 14.9 across a single independent test, Qwen is competitive in performance but should be considered alongside models with higher scores. Its pricing stands at $0.15 per million input tokens and $0.6 per million output tokens, making it cost-effective for applications that require frequent interactions or large-scale data processing.
Benchmark results
Independent, published benchmarks. Blended score 14.9 across 1 benchmark, last refreshed 2026-08-13. How scoring works →
| Benchmark | Measures | Score |
|---|---|---|
| AI Index | broad capability composite | 16.3 |
- Model ID
- qwen/qwen3-vl-30b-a3b-instruct
- Vendor
- qwen
- Released
- October 2025
- Tokenizer
- Qwen3
- Input Modalities
- text, image
- Output Modalities
- text
- Max Output
- 16,384 tokens
- Tool Calling
- ✓ supported
- Structured Output
- ✓ supported
- Reasoning Mode
- not supported
- Vision
- ✓ accepts images
- Audio
- no
- Moderated
- no
What it costs in practice
Computed from the current $0.15/M input and $0.60/M output rates. Run your own numbers →
| Job | Tokens | Cost |
|---|---|---|
| Summarize a 50-page report | 30k in / 1.5k out | under $0.01 |
| Classify 1,000 customer emails | 500k in / 50k out | $0.10 |
| A month of a busy support chatbot | 5M in / 2M out | $1.95 |
Price & spec history
Tracked daily by PicksByModel since 2026-07-17.
| Date | Input /M | Output /M | Context |
|---|---|---|---|
| 2026-08-04 | $0.15 | $0.60 | 262,144 |
| 2026-08-01 | $0.13 | $0.52 | 262,144 |
| 2026-07-23 | $0.15 | $0.60 | 262,144 |
| 2026-07-17 | $0.13 | $0.52 | 262,144 |
Similar models
Qwen: Qwen3 VL 8B Instruct
Qwen: Qwen3 VL 235B A22B Instruct
Qwen: Qwen3 Next 80B A3B Thinking
Qwen: Qwen Plus 0728 (thinking)
Qwen: Qwen3.8 Max
Qwen: Qwen3.7 Flash
Quick answers
- How much does Qwen: Qwen3 VL 30B A3B Instruct cost?
- $0.15 per million input tokens and $0.60 per million output tokens.
- What is Qwen: Qwen3 VL 30B A3B Instruct's context window?
- 262,144 tokens, roughly 393 pages of text in a single request.
- Does Qwen: Qwen3 VL 30B A3B Instruct support tool calling?
- It supports tool calling, structured output.
- Can Qwen: Qwen3 VL 30B A3B Instruct process images?
- Yes, it accepts image input.