Z.ai: GLM 4.5
by Z.aiZ.ai: GLM 4.5 is a large language model from Z.ai. It costs $0.600 per million input tokens and $2.20 per million output tokens. Its context window is 131K tokens.
- Input / 1M tokens
- $0.600
- Output / 1M tokens
- $2.20
- Cached input / 1M
- $0.110
- Context window
- 131K
On repeated prefixes
tokens
Who serves it cheapest
1 hosts serve GLM 4.5. Same weights, same API - the price difference is pure margin and routing.
| Provider | Input / 1M | Output / 1M | Context | Throughput | Uptime 24h |
|---|---|---|---|---|---|
| Z.AICheapestfp8 | $0.600 | $2.20 | 131K | - | 100.0% |
About GLM 4.5
GLM-4.5 is our latest flagship foundation model, purpose-built for agent-based applications. It leverages a Mixture-of-Experts (MoE) architecture and supports a context length of up to 128k tokens. GLM-4.5 delivers significantly...
Specifications
| Model ID | z-ai/glm-4.5 |
|---|---|
| Provider | Z.ai |
| Context window | 131K tokens |
| Max output | 98K tokens |
| Input modalities | text |
| Output modalities | text |
| Knowledge cutoff | 2024-12-31 |
| Open weights | Yes - zai-org/GLM-4.5 |
| Released | July 25, 2025 |
Cheaper alternatives
Models that cost less than GLM 4.5 while keeping at least half its context window and every input modality it supports.
Frequently asked
How much does Z.ai: GLM 4.5 cost?
$0.600 per million input tokens and $2.20 per million output tokens. Cached input reads cost $0.110 per million tokens.
What is the context window of Z.ai: GLM 4.5?
131K tokens, with up to 98K tokens of output per request.
Confirm against the source: Z.ai official pricing.