GLM API pricing

Every GLM model from Zhipu AI we track, with input, cached input and output cost in USD per 1M tokens. Zhipu publishes several genuinely free Flash models alongside paid GLM tiers, which is unusual among hosted APIs.

20 models · updated 2026-08-12 · prices in USD per 1,000,000 tokens

20
Models tracked
Free
Cheapest input — GLM-4.7-Flash
$2.20
Priciest input — GLM-4.5-X
1M
Largest context — GLM-5.2

Listed in the order Zhipu AI (GLM) publishes them.

GLM models with input, cached input and output pricing per million tokens
GLM-5.2$1.40$0.260$4.401M
GLM-5.1$1.40$0.260$4.40205K
GLM-5$1.00$0.200$3.20205K
GLM-5-Turbo$1.20$0.240$4.00203K
GLM-4.7$0.600$0.110$2.20205K
GLM-4.7-FlashX$0.070$0.010$0.400
GLM-4.6$0.600$0.110$2.20205K
GLM-4.5$0.600$0.110$2.20131K
GLM-4.5-X$2.20$0.450$8.90
GLM-4.5-Air$0.200$0.030$1.10131K
GLM-4.5-AirX$1.10$0.220$4.50
GLM-4-32B-0414-128K$0.100$0.100
GLM-4.7-FlashFreeFreeFree203K
GLM-4.5-FlashFreeFreeFree
GLM-5V-Turbo$1.20$0.240$4.00203K
GLM-4.6V$0.300$0.050$0.900131K
GLM-OCR$0.030$0.030
GLM-4.6V-FlashX$0.040$0.004$0.400
GLM-4.5V$0.600$0.110$1.8066K
GLM-4.6V-FlashFreeFreeFree

GLM pricing questions

How much does the GLM API cost?
GLM pricing runs from free to $2.20 per 1M input tokens depending on the model. GLM-4.5-X costs $2.20 per 1M input and $8.90 per 1M output tokens.
What is the cheapest GLM model?
GLM-4.7-Flash is the cheapest GLM model we track, at free per 1M input tokens and free per 1M output tokens.
Does GLM charge less for cached tokens?
Yes. Cached input is billed at a lower rate than fresh input — for GLM-5.2 it is $0.260 per 1M versus $1.40, about 81% cheaper. Caching applies to repeated prompt prefixes.
How often is this GLM pricing updated?
Every day. Prices are read directly from the published pricing page, normalized to USD per 1,000,000 tokens, and only recorded when a number actually changes — so the history shows real price movements rather than daily noise.

Compare with other providers