qwen

Qwen: Qwen3 VL 8B Instruct

Qwen3 VL 8B Instruct is an AI model that handles input from both images and text, with a maximum context length of 262,144 tokens. It supports tools but does not support complex reasoning or structured output. At $0.117 per million input tokens and $0.455 per million output tokens, Qwen3 VL 8B Instruct may be worth considering for users who need to process large volumes of data in various formats, particularly those working with images. With a blended benchmark score of 12.2 across one independent benchmark, this model stands out within its coverage area; however, more information is needed to make a definitive assessment of its performance relative to other models.

Quality Score
99/100
price + capability + benchmarks
Input Price
$0.12
per 1M tokens · checked 2026-09-04
Output Price
$0.46
per 1M tokens · checked 2026-09-04
Context Window
262,144
tokens

Benchmark results

Independent, published benchmarks. Blended score 12.2 across 1 benchmark, last refreshed 2026-09-04. How scoring works →

BenchmarkMeasuresScore
AI Index broad capability composite 13.6
Model ID
qwen/qwen3-vl-8b-instruct
Vendor
qwen
Released
October 2025
Tokenizer
Qwen3
Input Modalities
image, text
Output Modalities
text
Max Output
32,768 tokens
Tool Calling
✓ supported
Structured Output
✓ supported
Reasoning Mode
not supported
Vision
✓ accepts images
Audio
no
Moderated
no

What it costs in practice

Computed from the current $0.12/M input and $0.46/M output rates. Run your own numbers →

JobTokensCost
Summarize a 50-page report 30k in / 1.5k out under $0.01
Classify 1,000 customer emails 500k in / 50k out $0.08
A month of a busy support chatbot 5M in / 2M out $1.50

Price & spec history

Tracked daily by PicksByModel since 2026-07-17.

DateInput /MOutput /MContext
2026-07-22 $0.12 $0.46 262,144
2026-07-17 $0.12 $0.46 256,000

Similar models

Quick answers

How much does Qwen: Qwen3 VL 8B Instruct cost?
$0.12 per million input tokens and $0.46 per million output tokens.
What is Qwen: Qwen3 VL 8B Instruct's context window?
262,144 tokens, roughly 393 pages of text in a single request.
Does Qwen: Qwen3 VL 8B Instruct support tool calling?
It supports tool calling, structured output.
Can Qwen: Qwen3 VL 8B Instruct process images?
Yes, it accepts image input.