GLM-4.6V pricing
GLM-4.6V from Zhipu AI (GLM) costs $0.300 per 1M input tokens and $0.900 per 1M output tokens. Cached input is $0.050 per 1M. Context window 131,072 tokens.
API id glm-4.6v · Zhipu AI (GLM) · source · updated 2026-08-12
$0.300
Input / 1M tokens
$0.900
Output / 1M tokens
$0.050
Cached input / 1M
131K
Context window
What GLM-4.6V costs in practice
| Workload | Input tokens | Output tokens | Estimated cost |
|---|---|---|---|
| Short chat turn | 1,000 | 500 | $0.0008 |
| Document summary | 50,000 | 2,000 | $0.02 |
| 1M tokens in, 200K out | 1,000,000 | 200,000 | $0.48 |
Estimates use the standard tier and ignore caching discounts, so real bills with repeated prompt prefixes are typically lower.
Common questions
- How much does GLM-4.6V cost per 1M tokens?
- GLM-4.6V costs $0.300 per 1M input tokens and $0.900 per 1M output tokens on the standard tier. Cached input is $0.050 per 1M.
- What is GLM-4.6V's context window?
- GLM-4.6V accepts up to 131,072 tokens in a single request, and can return up to 32,768 output tokens.
- Is there a cheaper GLM model than GLM-4.6V?
- Yes. GLM-4.6V-Flash is free per 1M input tokens versus $0.300 for GLM-4.6V.
- Where does this GLM-4.6V price come from?
- Directly from Zhipu AI (GLM)'s published pricing page, re-read every day and normalized to USD per 1,000,000 tokens.