qwen

Qwen: Qwen3 VL 30B A3B Instruct

Qwen3 VL 30B A3B Instruct is capable of processing both text and images, with a context length of 262,144 tokens and support for external tools. It does not offer reasoning or structured output capabilities; however, it excels in handling multi-modal inputs effectively. This model is suitable for tasks that require understanding and processing of diverse data types, such as image captioning or text-image alignment. Given its blended benchmark score of 14.9 across a single independent test, Qwen is competitive in performance but should be considered alongside models with higher scores. Its pricing stands at $0.15 per million input tokens and $0.6 per million output tokens, making it cost-effective for applications that require frequent interactions or large-scale data processing.

Quality Score
99/100
price + capability + benchmarks
Input Price
$0.15
per 1M tokens
Output Price
$0.60
per 1M tokens
Context Window
262,144
tokens

Benchmark results

Independent, published benchmarks. Blended score 14.9 across 1 benchmark, last refreshed 2026-08-13. How scoring works →

BenchmarkMeasuresScore
AI Index broad capability composite 16.3
Model ID
qwen/qwen3-vl-30b-a3b-instruct
Vendor
qwen
Released
October 2025
Tokenizer
Qwen3
Input Modalities
text, image
Output Modalities
text
Max Output
16,384 tokens
Tool Calling
✓ supported
Structured Output
✓ supported
Reasoning Mode
not supported
Vision
✓ accepts images
Audio
no
Moderated
no

What it costs in practice

Computed from the current $0.15/M input and $0.60/M output rates. Run your own numbers →

JobTokensCost
Summarize a 50-page report 30k in / 1.5k out under $0.01
Classify 1,000 customer emails 500k in / 50k out $0.10
A month of a busy support chatbot 5M in / 2M out $1.95

Price & spec history

Tracked daily by PicksByModel since 2026-07-17.

DateInput /MOutput /MContext
2026-08-04 $0.15 $0.60 262,144
2026-08-01 $0.13 $0.52 262,144
2026-07-23 $0.15 $0.60 262,144
2026-07-17 $0.13 $0.52 262,144

Similar models

Quick answers

How much does Qwen: Qwen3 VL 30B A3B Instruct cost?
$0.15 per million input tokens and $0.60 per million output tokens.
What is Qwen: Qwen3 VL 30B A3B Instruct's context window?
262,144 tokens, roughly 393 pages of text in a single request.
Does Qwen: Qwen3 VL 30B A3B Instruct support tool calling?
It supports tool calling, structured output.
Can Qwen: Qwen3 VL 30B A3B Instruct process images?
Yes, it accepts image input.