Qwen2.5 VL 72B Instruct pricing

Qwen2.5 VL 72B Instruct from Alibaba (Qwen) costs $0.250 per 1M input tokens and $0.750 per 1M output tokens. Context window 128,000 tokens.

API id qwen2.5-vl-72b-instruct · Alibaba (Qwen) · source · updated 2026-08-11

$0.250
Input / 1M tokens
$0.750
Output / 1M tokens
Cached input / 1M
128K
Context window

What Qwen2.5 VL 72B Instruct costs in practice

Estimated cost at common usage volumes
WorkloadInput tokensOutput tokensEstimated cost
Short chat turn1,000500$0.0006
Document summary50,0002,000$0.01
1M tokens in, 200K out1,000,000200,000$0.40

Estimates use the standard tier and ignore caching discounts, so real bills with repeated prompt prefixes are typically lower.

Common questions

How much does Qwen2.5 VL 72B Instruct cost per 1M tokens?
Qwen2.5 VL 72B Instruct costs $0.250 per 1M input tokens and $0.750 per 1M output tokens on the standard tier.
What is Qwen2.5 VL 72B Instruct's context window?
Qwen2.5 VL 72B Instruct accepts up to 128,000 tokens in a single request.
Is there a cheaper Qwen model than Qwen2.5 VL 72B Instruct?
Yes. Qwen2.5 7B Instruct is $0.100 per 1M input tokens versus $0.250 for Qwen2.5 VL 72B Instruct, with a 32,768 token context window.
Where does this Qwen2.5 VL 72B Instruct price come from?
From the OpenRouter catalogue, because Alibaba (Qwen) does not publish machine-readable pricing. OpenRouter is a reseller, so its rate can differ from Alibaba (Qwen)'s own — verify before committing spend.

Other Qwen models

See all Qwen pricing →