Qwen: Qwen3 VL 8B Instruct
Qwen3 VL 8B Instruct is an AI model that handles input from both images and text, with a maximum context length of 262,144 tokens. It supports tools but does not support complex reasoning or structured output. At $0.117 per million input tokens and $0.455 per million output tokens, Qwen3 VL 8B Instruct may be worth considering for users who need to process large volumes of data in various formats, particularly those working with images. With a blended benchmark score of 12.2 across one independent benchmark, this model stands out within its coverage area; however, more information is needed to make a definitive assessment of its performance relative to other models.
Benchmark results
Independent, published benchmarks. Blended score 12.2 across 1 benchmark, last refreshed 2026-09-04. How scoring works →
| Benchmark | Measures | Score |
|---|---|---|
| AI Index | broad capability composite | 13.6 |
- Model ID
- qwen/qwen3-vl-8b-instruct
- Vendor
- qwen
- Released
- October 2025
- Tokenizer
- Qwen3
- Input Modalities
- image, text
- Output Modalities
- text
- Max Output
- 32,768 tokens
- Tool Calling
- ✓ supported
- Structured Output
- ✓ supported
- Reasoning Mode
- not supported
- Vision
- ✓ accepts images
- Audio
- no
- Moderated
- no
What it costs in practice
Computed from the current $0.12/M input and $0.46/M output rates. Run your own numbers →
| Job | Tokens | Cost |
|---|---|---|
| Summarize a 50-page report | 30k in / 1.5k out | under $0.01 |
| Classify 1,000 customer emails | 500k in / 50k out | $0.08 |
| A month of a busy support chatbot | 5M in / 2M out | $1.50 |
Price & spec history
Tracked daily by PicksByModel since 2026-07-17.
| Date | Input /M | Output /M | Context |
|---|---|---|---|
| 2026-07-22 | $0.12 | $0.46 | 262,144 |
| 2026-07-17 | $0.12 | $0.46 | 256,000 |
Similar models
Qwen: Qwen3 VL 30B A3B Instruct
Qwen: Qwen3.8 Flash
Qwen: Qwen3.8 27B
Qwen: Qwen3.8 Max
Qwen: Qwen3.7 Flash
Qwen: Qwen3.7 Plus
Quick answers
- How much does Qwen: Qwen3 VL 8B Instruct cost?
- $0.12 per million input tokens and $0.46 per million output tokens.
- What is Qwen: Qwen3 VL 8B Instruct's context window?
- 262,144 tokens, roughly 393 pages of text in a single request.
- Does Qwen: Qwen3 VL 8B Instruct support tool calling?
- It supports tool calling, structured output.
- Can Qwen: Qwen3 VL 8B Instruct process images?
- Yes, it accepts image input.