Qwen2.5 72B Instruct pricing

Qwen2.5 72B Instruct from Alibaba (Qwen) costs $0.360 per 1M input tokens and $0.400 per 1M output tokens. Context window 32,768 tokens.

API id qwen-2.5-72b-instruct · Alibaba (Qwen) · source · updated 2026-08-11

$0.360
Input / 1M tokens
$0.400
Output / 1M tokens
Cached input / 1M
33K
Context window

What Qwen2.5 72B Instruct costs in practice

Estimated cost at common usage volumes
WorkloadInput tokensOutput tokensEstimated cost
Short chat turn1,000500$0.0006
Document summary50,0002,000$0.02
1M tokens in, 200K out1,000,000200,000$0.44

Estimates use the standard tier and ignore caching discounts, so real bills with repeated prompt prefixes are typically lower.

Common questions

How much does Qwen2.5 72B Instruct cost per 1M tokens?
Qwen2.5 72B Instruct costs $0.360 per 1M input tokens and $0.400 per 1M output tokens on the standard tier.
What is Qwen2.5 72B Instruct's context window?
Qwen2.5 72B Instruct accepts up to 32,768 tokens in a single request, and can return up to 16,384 output tokens.
Is there a cheaper Qwen model than Qwen2.5 72B Instruct?
Yes. Qwen2.5 7B Instruct is $0.100 per 1M input tokens versus $0.360 for Qwen2.5 72B Instruct, with a 32,768 token context window.
Where does this Qwen2.5 72B Instruct price come from?
From the OpenRouter catalogue, because Alibaba (Qwen) does not publish machine-readable pricing. OpenRouter is a reseller, so its rate can differ from Alibaba (Qwen)'s own — verify before committing spend.

Other Qwen models

See all Qwen pricing →