Qwen: Qwen3 VL 30B A3B Instruct
Qwen3 VL 30B A3B Instruct is a versatile AI model that supports both text and image inputs within a substantial context length of 262,144 tokens. It also offers tools support but lacks reasoning capabilities or structured output options. The benchmark score of 15.0 across one independent test stands at 16.5 on the AI index, last updated on July 31, 2026. This model is suitable for users needing a robust text and image processing solution with tools support but who prioritize cost-effectiveness given its pricing at $0.15 per million input tokens and $0.6 per million output tokens. Its benchmark standing, while not extensively covered, places it competitively in the market.
Benchmark results
Independent, published benchmarks. Blended score 15.0 across 1 benchmark, last refreshed 2026-07-31. How scoring works →
| Benchmark | Measures | Score |
|---|---|---|
| AI Index | broad capability composite | 16.5 |
- Model ID
- qwen/qwen3-vl-30b-a3b-instruct
- Vendor
- qwen
- Released
- October 2025
- Tokenizer
- Qwen3
- Input Modalities
- text, image
- Output Modalities
- text
- Max Output
- 16,384 tokens
- Tool Calling
- ✓ supported
- Structured Output
- ✓ supported
- Reasoning Mode
- not supported
- Vision
- ✓ accepts images
- Audio
- no
- Moderated
- no
What it costs in practice
Computed from the current $0.15/M input and $0.60/M output rates. Run your own numbers →
| Job | Tokens | Cost |
|---|---|---|
| Summarize a 50-page report | 30k in / 1.5k out | under $0.01 |
| Classify 1,000 customer emails | 500k in / 50k out | $0.10 |
| A month of a busy support chatbot | 5M in / 2M out | $1.95 |
Price & spec history
Tracked daily by PicksByModel since 2026-07-17.
| Date | Input /M | Output /M | Context |
|---|---|---|---|
| 2026-07-23 | $0.15 | $0.60 | 262,144 |
| 2026-07-17 | $0.13 | $0.52 | 262,144 |
Similar models
Qwen: Qwen3 VL 8B Instruct
Qwen: Qwen3 Next 80B A3B Thinking
Qwen: Qwen Plus 0728 (thinking)
Qwen: Qwen3.7 Flash
Qwen: Qwen3.7 Plus
Qwen: Qwen3.5 Plus 2026-04-20
Quick answers
- How much does Qwen: Qwen3 VL 30B A3B Instruct cost?
- $0.15 per million input tokens and $0.60 per million output tokens.
- What is Qwen: Qwen3 VL 30B A3B Instruct's context window?
- 262,144 tokens, roughly 393 pages of text in a single request.
- Does Qwen: Qwen3 VL 30B A3B Instruct support tool calling?
- It supports tool calling, structured output.
- Can Qwen: Qwen3 VL 30B A3B Instruct process images?
- Yes, it accepts image input.