Skip to content
LLMs
Z.ai logo

Z.ai: GLM 5 Turbo

by Z.ai

Z.ai: GLM 5 Turbo is a large language model from Z.ai. It costs $1.20 per million input tokens and $4.00 per million output tokens. Its context window is 203K tokens.

ReasoningTool callingStructured outputPrompt caching
Input / 1M tokens
$1.20
Output / 1M tokens
$4.00
Cached input / 1M
$0.240

On repeated prefixes

Context window
203K

tokens

Who serves it cheapest

1 hosts serve GLM 5 Turbo. Same weights, same API - the price difference is pure margin and routing.

Providers serving Z.ai: GLM 5 Turbo, cheapest first
ProviderInput / 1MOutput / 1MContextThroughputUptime 24h
Z.AICheapest$1.20$4.00203K-100.0%

About GLM 5 Turbo

GLM-5 Turbo is a new model from Z.ai designed for fast inference and strong performance in agent-driven environments such as OpenClaw scenarios. It is deeply optimized for real-world agent workflows...

Specifications

Z.ai: GLM 5 Turbo specifications
Model IDz-ai/glm-5-turbo
ProviderZ.ai
Context window203K tokens
Max output131K tokens
Input modalitiestext
Output modalitiestext
Knowledge cutoff-
Open weightsNo
ReleasedMarch 15, 2026

Cheaper alternatives

Models that cost less than GLM 5 Turbo while keeping at least half its context window and every input modality it supports.

Frequently asked

How much does Z.ai: GLM 5 Turbo cost?

$1.20 per million input tokens and $4.00 per million output tokens. Cached input reads cost $0.240 per million tokens.

What is the context window of Z.ai: GLM 5 Turbo?

203K tokens, with up to 131K tokens of output per request.

Confirm against the source: Z.ai official pricing.