Skip to content
LLMs
Z.ai logo

Z.ai: GLM 4.5

by Z.ai

Z.ai: GLM 4.5 is a large language model from Z.ai. It costs $0.600 per million input tokens and $2.20 per million output tokens. Its context window is 131K tokens.

ReasoningTool callingStructured outputPrompt cachingOpen weights
Input / 1M tokens
$0.600
Output / 1M tokens
$2.20
Cached input / 1M
$0.110

On repeated prefixes

Context window
131K

tokens

Who serves it cheapest

1 hosts serve GLM 4.5. Same weights, same API - the price difference is pure margin and routing.

Providers serving Z.ai: GLM 4.5, cheapest first
ProviderInput / 1MOutput / 1MContextThroughputUptime 24h
Z.AICheapestfp8$0.600$2.20131K-100.0%

About GLM 4.5

GLM-4.5 is our latest flagship foundation model, purpose-built for agent-based applications. It leverages a Mixture-of-Experts (MoE) architecture and supports a context length of up to 128k tokens. GLM-4.5 delivers significantly...

Specifications

Z.ai: GLM 4.5 specifications
Model IDz-ai/glm-4.5
ProviderZ.ai
Context window131K tokens
Max output98K tokens
Input modalitiestext
Output modalitiestext
Knowledge cutoff2024-12-31
Open weightsYes - zai-org/GLM-4.5
ReleasedJuly 25, 2025

Cheaper alternatives

Models that cost less than GLM 4.5 while keeping at least half its context window and every input modality it supports.

Frequently asked

How much does Z.ai: GLM 4.5 cost?

$0.600 per million input tokens and $2.20 per million output tokens. Cached input reads cost $0.110 per million tokens.

What is the context window of Z.ai: GLM 4.5?

131K tokens, with up to 98K tokens of output per request.

Confirm against the source: Z.ai official pricing.