GLM-4.7-Flash pricing

GLM-4.7-Flash from Zhipu AI (GLM) costs free per 1M input tokens and free per 1M output tokens. Cached input is free per 1M. Context window 202,752 tokens.

API id glm-4.7-flash · Zhipu AI (GLM) · source · updated 2026-08-11

Free
Input / 1M tokens
Free
Output / 1M tokens
Free
Cached input / 1M
203K
Context window

What GLM-4.7-Flash costs in practice

Estimated cost at common usage volumes
WorkloadInput tokensOutput tokensEstimated cost
Short chat turn1,000500Free
Document summary50,0002,000Free
1M tokens in, 200K out1,000,000200,000Free

Estimates use the standard tier and ignore caching discounts, so real bills with repeated prompt prefixes are typically lower.

Common questions

How much does GLM-4.7-Flash cost per 1M tokens?
GLM-4.7-Flash costs free per 1M input tokens and free per 1M output tokens on the standard tier. Cached input is free per 1M.
What is GLM-4.7-Flash's context window?
GLM-4.7-Flash accepts up to 202,752 tokens in a single request, and can return up to 16,384 output tokens.
Where does this GLM-4.7-Flash price come from?
Directly from Zhipu AI (GLM)'s published pricing page, re-read every day and normalized to USD per 1,000,000 tokens.

Other GLM models

See all GLM pricing →