qwen

Qwen: Qwen3.8 Flash

Qwen3.8 Flash is an AI model capable of handling text, images, and video inputs, with a context length of one million tokens and support for reasoning tasks. This versatile model also supports tools integration but does not offer structured output functionality. While Qwen3.8 Flash excels in its broad input modalities and reasoning capabilities, users should note that it lacks formal benchmarking coverage, meaning its performance relative to other models remains unproven. Given its pricing at $0.16 per million input tokens and $0.47 per million output tokens, this model is best suited for organizations requiring a robust multimodal processing capability within their budget constraints.

Quality Score
100/100
price + capability + benchmarks
Input Price
$0.16
per 1M tokens · checked 2026-08-27
Output Price
$0.47
per 1M tokens · checked 2026-08-27
Context Window
1,000,000
tokens
Model ID
qwen/qwen3.8-flash
Vendor
qwen
Released
August 2026
Tokenizer
Qwen
Input Modalities
text, image, video
Output Modalities
text
Max Output
131,072 tokens
Tool Calling
✓ supported
Structured Output
✓ supported
Reasoning Mode
✓ supported
Vision
✓ accepts images
Audio
no
Moderated
no

What it costs in practice

Computed from the current $0.16/M input and $0.47/M output rates. Run your own numbers →

JobTokensCost
Summarize a 50-page report 30k in / 1.5k out under $0.01
Classify 1,000 customer emails 500k in / 50k out $0.10
A month of a busy support chatbot 5M in / 2M out $1.74

Strong choice for

Category rankings

Where Qwen: Qwen3.8 Flash places across the 5 categories it ranks in. How we rank →

#CategoryScore
#1 Video Auto-TaggingVideo · of 25 ranked 123
#12 Social Media PostsWriting · of 25 ranked 119
#12 Voice Assistant BackendVoice · of 25 ranked 123
#22 Self-Hosted / LocalCost · of 25 ranked 117
#23 Real-Time ChatLatency · of 25 ranked 117

Similar models

Quick answers

How much does Qwen: Qwen3.8 Flash cost?
$0.16 per million input tokens and $0.47 per million output tokens.
What is Qwen: Qwen3.8 Flash's context window?
1,000,000 tokens, roughly 1,500 pages of text in a single request.
Does Qwen: Qwen3.8 Flash support tool calling?
It supports tool calling, structured output, a reasoning mode.
Can Qwen: Qwen3.8 Flash process images?
Yes, it accepts image input.