z-ai

Z.ai: GLM 4.5V

The Z.ai: GLM 4.5V is a multimodal AI model that can process both text and image inputs. It has a large context length of 65536 tokens and supports reasoning capabilities. However, it does not have structured output and its open weights status is unknown. For users with a budget-friendly approach in mind, the Z.ai: GLM 4.5V might be worth considering due to its relatively low input price point of $0.6 per million tokens. Its blended benchmark score of 5.9 also indicates a decent performance across one independent benchmark, but further testing is needed for a more comprehensive understanding of its capabilities.

Quality Score
87/100
price + capability + benchmarks
Input Price
$0.60
per 1M tokens · checked 2026-09-08
Output Price
$1.80
per 1M tokens · checked 2026-09-08
Context Window
65,536
tokens

Benchmark results

Independent, published benchmarks. Blended score 5.9 across 1 benchmark, last refreshed 2026-09-08. How scoring works →

BenchmarkMeasuresScore
AI Index broad capability composite 11.0
Model ID
z-ai/glm-4.5v
Vendor
z-ai
Released
August 2025
Tokenizer
Other
Input Modalities
text, image
Output Modalities
text
Max Output
16,384 tokens
Tool Calling
✓ supported
Structured Output
✓ supported
Reasoning Mode
✓ supported
Vision
✓ accepts images
Audio
no
Moderated
no

What it costs in practice

Computed from the current $0.60/M input and $1.80/M output rates. Run your own numbers →

JobTokensCost
Summarize a 50-page report 30k in / 1.5k out $0.02
Classify 1,000 customer emails 500k in / 50k out $0.39
A month of a busy support chatbot 5M in / 2M out $6.60

Similar models

Quick answers

How much does Z.ai: GLM 4.5V cost?
$0.60 per million input tokens and $1.80 per million output tokens.
What is Z.ai: GLM 4.5V's context window?
65,536 tokens, roughly 98 pages of text in a single request.
Does Z.ai: GLM 4.5V support tool calling?
It supports tool calling, structured output, a reasoning mode.
Can Z.ai: GLM 4.5V process images?
Yes, it accepts image input.