GLM-4.6V pricing

GLM-4.6V from Zhipu AI (GLM) costs $0.300 per 1M input tokens and $0.900 per 1M output tokens. Cached input is $0.050 per 1M. Context window 131,072 tokens.

API id glm-4.6v · Zhipu AI (GLM) · source · updated 2026-08-12

$0.300
Input / 1M tokens
$0.900
Output / 1M tokens
$0.050
Cached input / 1M
131K
Context window

What GLM-4.6V costs in practice

Estimated cost at common usage volumes
WorkloadInput tokensOutput tokensEstimated cost
Short chat turn1,000500$0.0008
Document summary50,0002,000$0.02
1M tokens in, 200K out1,000,000200,000$0.48

Estimates use the standard tier and ignore caching discounts, so real bills with repeated prompt prefixes are typically lower.

Common questions

How much does GLM-4.6V cost per 1M tokens?
GLM-4.6V costs $0.300 per 1M input tokens and $0.900 per 1M output tokens on the standard tier. Cached input is $0.050 per 1M.
What is GLM-4.6V's context window?
GLM-4.6V accepts up to 131,072 tokens in a single request, and can return up to 32,768 output tokens.
Is there a cheaper GLM model than GLM-4.6V?
Yes. GLM-4.6V-Flash is free per 1M input tokens versus $0.300 for GLM-4.6V.
Where does this GLM-4.6V price come from?
Directly from Zhipu AI (GLM)'s published pricing page, re-read every day and normalized to USD per 1,000,000 tokens.

Other GLM models

See all GLM pricing →