Qwen2.5 72B Instruct pricing
Qwen2.5 72B Instruct from Alibaba (Qwen) costs $0.360 per 1M input tokens and $0.400 per 1M output tokens. Context window 32,768 tokens.
API id qwen-2.5-72b-instruct · Alibaba (Qwen) · source · updated 2026-08-11
$0.360
Input / 1M tokens
$0.400
Output / 1M tokens
—
Cached input / 1M
33K
Context window
What Qwen2.5 72B Instruct costs in practice
| Workload | Input tokens | Output tokens | Estimated cost |
|---|---|---|---|
| Short chat turn | 1,000 | 500 | $0.0006 |
| Document summary | 50,000 | 2,000 | $0.02 |
| 1M tokens in, 200K out | 1,000,000 | 200,000 | $0.44 |
Estimates use the standard tier and ignore caching discounts, so real bills with repeated prompt prefixes are typically lower.
Common questions
- How much does Qwen2.5 72B Instruct cost per 1M tokens?
- Qwen2.5 72B Instruct costs $0.360 per 1M input tokens and $0.400 per 1M output tokens on the standard tier.
- What is Qwen2.5 72B Instruct's context window?
- Qwen2.5 72B Instruct accepts up to 32,768 tokens in a single request, and can return up to 16,384 output tokens.
- Is there a cheaper Qwen model than Qwen2.5 72B Instruct?
- Yes. Qwen2.5 7B Instruct is $0.100 per 1M input tokens versus $0.360 for Qwen2.5 72B Instruct, with a 32,768 token context window.
- Where does this Qwen2.5 72B Instruct price come from?
- From the OpenRouter catalogue, because Alibaba (Qwen) does not publish machine-readable pricing. OpenRouter is a reseller, so its rate can differ from Alibaba (Qwen)'s own — verify before committing spend.
Other Qwen models
- Qwen3.8 Max $2.00
- Qwen3.7 Flash $0.030
- Qwen3.7 Plus $0.320
- Qwen3.7 Max $1.48
- Qwen3.5 Plus 2026-04-20 $0.300
- Qwen3.6 Flash $0.188
- Qwen3.6 35B A3B $0.150
- Qwen3.6 Max Preview $1.03
- Qwen3.6 27B $0.600
- Qwen3.6 Plus $0.325
- Qwen3.5-9B $0.100
- Qwen3.5-35B-A3B $0.140
- Qwen3.5-27B $0.195
- Qwen3.5-122B-A10B $0.290
- Qwen3.5-Flash $0.065
- Qwen3.5 Plus 2026-02-15 $0.260