GLM-5-Turbo pricing

GLM-5-Turbo from Zhipu AI (GLM) costs $1.20 per 1M input tokens and $4.00 per 1M output tokens. Cached input is $0.240 per 1M. Context window 202,752 tokens.

API id glm-5-turbo · Zhipu AI (GLM) · source · updated 2026-08-11

$1.20
Input / 1M tokens
$4.00
Output / 1M tokens
$0.240
Cached input / 1M
203K
Context window

What GLM-5-Turbo costs in practice

Estimated cost at common usage volumes
WorkloadInput tokensOutput tokensEstimated cost
Short chat turn1,000500$0.0032
Document summary50,0002,000$0.07
1M tokens in, 200K out1,000,000200,000$2.00

Estimates use the standard tier and ignore caching discounts, so real bills with repeated prompt prefixes are typically lower.

Common questions

How much does GLM-5-Turbo cost per 1M tokens?
GLM-5-Turbo costs $1.20 per 1M input tokens and $4.00 per 1M output tokens on the standard tier. Cached input is $0.240 per 1M.
What is GLM-5-Turbo's context window?
GLM-5-Turbo accepts up to 202,752 tokens in a single request, and can return up to 131,072 output tokens.
Is there a cheaper GLM model than GLM-5-Turbo?
Yes. GLM-4.6V-Flash is free per 1M input tokens versus $1.20 for GLM-5-Turbo.
Where does this GLM-5-Turbo price come from?
Directly from Zhipu AI (GLM)'s published pricing page, re-read every day and normalized to USD per 1,000,000 tokens.

Other GLM models

See all GLM pricing →