Z.ai: GLM 4.6V
Z.ai: GLM 4.6V is a powerful AI model from z-ai that can handle text, images, and videos, with a context length of up to 131,072 tokens and the ability to support reasoning. It offers structured output capabilities and integrates tools for enhanced functionality. This model should be considered by users requiring versatility across multiple input modalities and robust reasoning abilities, despite its premium pricing at $0.3 per million input tokens and $0.9 per million output tokens. Its blended benchmark score of 16.8 places it favorably in performance evaluations, making it a strong contender for those who prioritize effectiveness in their AI solutions.
Benchmark results
Independent, published benchmarks. Blended score 16.8 across 1 benchmark, last refreshed 2026-07-25. How scoring works →
| Benchmark | Measures | Score |
|---|---|---|
| AI Index | broad capability composite | 18.1 |
- Model ID
- z-ai/glm-4.6v
- Vendor
- z-ai
- Released
- December 2025
- Tokenizer
- Other
- Input Modalities
- image, text, video
- Output Modalities
- text
- Max Output
- 32,768 tokens
- Tool Calling
- ✓ supported
- Structured Output
- ✓ supported
- Reasoning Mode
- ✓ supported
- Vision
- ✓ accepts images
- Audio
- no
- Moderated
- no
What it costs in practice
Computed from the current $0.30/M input and $0.90/M output rates. Run your own numbers →
| Job | Tokens | Cost |
|---|---|---|
| Summarize a 50-page report | 30k in / 1.5k out | $0.01 |
| Classify 1,000 customer emails | 500k in / 50k out | $0.20 |
| A month of a busy support chatbot | 5M in / 2M out | $3.30 |
Category rankings
Where Z.ai: GLM 4.6V places across the 3 categories it ranks in. How we rank →
| # | Category | Score |
|---|---|---|
| #19 | Social Media PostsWriting · of 25 ranked | 119 |
| #19 | Voice Assistant BackendVoice · of 25 ranked | 123 |
| #21 | Real-Time ChatLatency · of 25 ranked | 117 |
Similar models
Z.ai: GLM 5V Turbo
Z.ai: GLM 4.7 Flash
Z.ai: GLM 4.7
Z.ai: GLM 4.6
Z.ai: GLM 5.2
Z.ai: GLM 5
Quick answers
- How much does Z.ai: GLM 4.6V cost?
- $0.30 per million input tokens and $0.90 per million output tokens.
- What is Z.ai: GLM 4.6V's context window?
- 131,072 tokens, roughly 196 pages of text in a single request.
- Does Z.ai: GLM 4.6V support tool calling?
- It supports tool calling, structured output, a reasoning mode.
- Can Z.ai: GLM 4.6V process images?
- Yes, it accepts image input.