Fresh paint, wet floors — we're rebuilding CostOfToken into a cross-provider price comparison. Some corners are still under construction; the full version is coming up soon.

GLM-4.6V-Flash pricing

GLM-4.6V-Flash from Zhipu AI (GLM) costs free per 1M input tokens and free per 1M output tokens. Cached input is free per 1M.

API id glm-4.6v-flash · Zhipu AI (GLM) · source · updated 2026-09-25 · general-purpose

Free
Input / 1M tokens
Free
Output / 1M tokens
Free
Cached input / 1M
—
Context window

GLM-4.6V-Flash specifications

API id
glm-4.6v-flash
Model type
general-purpose
Context window
— tokens
Max output
not published
Currency
USD
Price source
First-party
Last checked
2026-09-25

Tags

  • fast

What GLM-4.6V-Flash costs in practice

Estimated cost at common usage volumes
WorkloadInput tokensOutput tokensEstimated cost
Short chat turn1,000500Free
Document summary50,0002,000Free
1M tokens in, 200K out1,000,000200,000Free

Estimates use the standard tier and ignore caching discounts, so real bills with repeated prompt prefixes are typically lower.

Common questions

How much does GLM-4.6V-Flash cost per 1M tokens?
GLM-4.6V-Flash costs free per 1M input tokens and free per 1M output tokens on the standard tier. Cached input is free per 1M.
What is GLM-4.6V-Flash's context window?
Zhipu AI (GLM) does not publish a context window for GLM-4.6V-Flash on its pricing page.
Where does this GLM-4.6V-Flash price come from?
Directly from Zhipu AI (GLM)'s published pricing page, re-read every day and normalized to USD per 1,000,000 tokens.

Head-to-head comparisons

Other GLM models

See all GLM pricing →