Qwen: Qwen3.5-9B
Qwen3.5-9B is a versatile AI model capable of processing text, images, and videos, with an impressive context length of 262,144 tokens. It supports both reasoning and the use of tools, enabling more complex tasks. Unlike some models that lack structured output capabilities, Qwen can handle a variety of inputs seamlessly. While Qwen excels in certain areas such as AI Index coding (scoring 47.3) and overall blended benchmarking (30.8), its performance in agentic tasks is relatively lower at 11.5. Given the pricing at $0.1 per input token and $0.15 per output token, it may be more cost-effective for applications requiring a mix of text, image, or video processing rather than scenarios demanding high agentic capabilities.
Benchmark results
Independent, published benchmarks. Blended score 30.8 across 3 benchmarks, last refreshed 2026-08-13. How scoring works →
| Benchmark | Measures | Score |
|---|---|---|
| AI Index | broad capability composite | 36.0 |
| AI Index Coding | software engineering tasks | 47.3 |
| AI Index Agentic | multi-step tool-using tasks | 11.5 |
- Model ID
- qwen/qwen3.5-9b
- Vendor
- qwen
- Released
- March 2026
- Tokenizer
- Qwen3
- Input Modalities
- text, image, video
- Output Modalities
- text
- Max Output
- 262,144 tokens
- Tool Calling
- ✓ supported
- Structured Output
- ✓ supported
- Reasoning Mode
- ✓ supported
- Vision
- ✓ accepts images
- Audio
- no
- Moderated
- no
What it costs in practice
Computed from the current $0.10/M input and $0.15/M output rates. Run your own numbers →
| Job | Tokens | Cost |
|---|---|---|
| Summarize a 50-page report | 30k in / 1.5k out | under $0.01 |
| Classify 1,000 customer emails | 500k in / 50k out | $0.06 |
| A month of a busy support chatbot | 5M in / 2M out | $0.80 |
Category rankings
Where Qwen: Qwen3.5-9B places across the 6 categories it ranks in. How we rank →
| # | Category | Score |
|---|---|---|
| #6 | Social Media PostsWriting · of 25 ranked | 120 |
| #6 | Voice Assistant BackendVoice · of 25 ranked | 124 |
| #6 | Cheap Bulk InferenceCost · of 25 ranked | 137 |
| #8 | Self-Hosted / LocalCost · of 25 ranked | 117 |
| #8 | Real-Time ChatLatency · of 25 ranked | 118 |
| #25 | Video Auto-TaggingVideo · of 25 ranked | 123 |
Similar models
Qwen: Qwen3.8 Max
Qwen: Qwen3.7 Flash
Qwen: Qwen3.7 Plus
Qwen: Qwen3.5 Plus 2026-04-20
Qwen: Qwen3.6 Flash
Qwen: Qwen3.6 35B A3B
Quick answers
- How much does Qwen: Qwen3.5-9B cost?
- $0.10 per million input tokens and $0.15 per million output tokens.
- What is Qwen: Qwen3.5-9B's context window?
- 262,144 tokens, roughly 393 pages of text in a single request.
- Does Qwen: Qwen3.5-9B support tool calling?
- It supports tool calling, structured output, a reasoning mode.
- Can Qwen: Qwen3.5-9B process images?
- Yes, it accepts image input.