qwen

Qwen: Qwen3 32B

This AI model, Qwen3 32B, is offered by vendor qwen and can handle text inputs up to 131072 tokens in length. It supports both text input modality and tools, with the ability to perform reasoning tasks but does not provide structured output. Given its price point of $0.08 per million input tokens and $0.28 per million output tokens, this model may be worth considering for those who value strong performance at a premium cost. Its benchmark standing is reflected in a blended score of 33.3 across three independent benchmarks, with notable scores on eqbench (46.4) and aider_polyglot (40.0).

Open weights 32.8B parameters, about 18.3 GiB at Q4_K_M: fits comfortably on 4 of 17 consumer graphics cards in the PicksByCard catalog. See which cards run it →
Quality Score
91/100
price + capability + benchmarks
Input Price
$0.08
per 1M tokens · checked 2026-09-10
Output Price
$0.28
per 1M tokens · checked 2026-09-10
Context Window
131,072
tokens

Benchmark results

Independent, published benchmarks. Blended score 33.3 across 3 benchmarks, last refreshed 2026-09-10. How scoring works →

BenchmarkMeasuresScore
Aider Polyglot code editing across languages 40.0
AI Index broad capability composite 12.1
EQ-Bench emotional understanding in dialogue 46.4
Model ID
qwen/qwen3-32b
Vendor
qwen
Released
April 2025
Tokenizer
Qwen3
Input Modalities
text
Output Modalities
text
Max Output
16,384 tokens
Tool Calling
✓ supported
Structured Output
✓ supported
Reasoning Mode
✓ supported
Vision
text only
Audio
no
Moderated
no

What it costs in practice

Computed from the current $0.08/M input and $0.28/M output rates. Run your own numbers →

JobTokensCost
Summarize a 50-page report 30k in / 1.5k out under $0.01
Classify 1,000 customer emails 500k in / 50k out $0.05
A month of a busy support chatbot 5M in / 2M out $0.96

Similar models

Quick answers

How much does Qwen: Qwen3 32B cost?
$0.08 per million input tokens and $0.28 per million output tokens.
What is Qwen: Qwen3 32B's context window?
131,072 tokens, roughly 196 pages of text in a single request.
Does Qwen: Qwen3 32B support tool calling?
It supports tool calling, structured output, a reasoning mode.
Can Qwen: Qwen3 32B process images?
No, it is text-only on the input side.