Z.ai: GLM Latest
Z.ai: GLM Latest is designed to handle text-based inputs with a vast context length of 1048576 tokens and supports both reasoning and tool integration. This model can process extensive information and make use of external tools for enhanced functionality. While Z.ai: GLM Latest offers robust features like support for reasoning and tool usage, its current benchmark coverage is limited. Given its pricing at $1.4 per million input tokens and $4.4 per million output tokens, this model might be more suitable for organizations that require a combination of extensive context handling and integrated tools, even though no independent benchmarks are available to assess its performance comprehensively.
- Model ID
- ~z-ai/glm-latest
- Vendor
- ~z-ai
- Released
- August 2026
- Tokenizer
- Router
- Input Modalities
- text
- Output Modalities
- text
- Max Output
- 131,072 tokens
- Tool Calling
- ✓ supported
- Structured Output
- ✓ supported
- Reasoning Mode
- ✓ supported
- Vision
- text only
- Audio
- no
- Moderated
- no
What it costs in practice
Computed from the current $1.40/M input and $4.40/M output rates. Run your own numbers →
| Job | Tokens | Cost |
|---|---|---|
| Summarize a 50-page report | 30k in / 1.5k out | $0.05 |
| Classify 1,000 customer emails | 500k in / 50k out | $0.92 |
| A month of a busy support chatbot | 5M in / 2M out | $15.80 |
Quick answers
- How much does Z.ai: GLM Latest cost?
- $1.40 per million input tokens and $4.40 per million output tokens.
- What is Z.ai: GLM Latest's context window?
- 1,048,576 tokens, roughly 1,572 pages of text in a single request.
- Does Z.ai: GLM Latest support tool calling?
- It supports tool calling, structured output, a reasoning mode.
- Can Z.ai: GLM Latest process images?
- No, it is text-only on the input side.