Qwen2.5 VL 72B Instruct pricing
Qwen2.5-VL is proficient in recognizing common objects such as flowers, birds, fish, and insects. It is also highly capable of analyzing texts, charts, icons, graphics, and layouts within images. Qwen2.5 VL 72B Instruct from Alibaba (Qwen) costs $0.800 per 1M input tokens and $1.00 per 1M output tokens. Cached input is $0.400 per 1M. Context window 128,000 tokens.
API id qwen2.5-vl-72b-instruct · Alibaba (Qwen) · source · updated 2026-09-25 · general-purpose
Qwen2.5 VL 72B Instruct specifications
- Provider
- Alibaba (Qwen)
- API id
qwen2.5-vl-72b-instruct- Model type
- general-purpose
- Context window
- 128K tokens
- Max output
- 115K tokens
- Currency
- USD
- Price source
- Provider API
- Last checked
- 2026-09-25
Tags
- vision
Qwen2.5 VL 72B Instruct across providers
The same model, priced by every seller we track. Ranked by the cost of 1M input + 1M output tokens, standard tier; savings measured against buying from the vendor directly.
| Provider | Input /1M | Cached /1M | Output /1M | 1M in + 1M out | vs vendor |
|---|---|---|---|---|---|
| $0.800 | $0.400 | $1.00 | $1.80 | — | |
OpenRouterHost | $0.800 | $0.400 | $1.00 | $1.80 | — |
What Qwen2.5 VL 72B Instruct costs in practice
| Workload | Input tokens | Output tokens | Estimated cost |
|---|---|---|---|
| Short chat turn | 1,000 | 500 | $0.0013 |
| Document summary | 50,000 | 2,000 | $0.04 |
| 1M tokens in, 200K out | 1,000,000 | 200,000 | $1.00 |
Estimates use the standard tier and ignore caching discounts, so real bills with repeated prompt prefixes are typically lower.
Qwen2.5 VL 72B Instruct price history
| Recorded | Input /1M | Output /1M |
|---|---|---|
| 2026-09-04 | $0.800 | $1.00 |
| 2026-08-26 | $0.250 | $0.750 |
| 2026-08-17 | $0.800 | $1.00 |
| 2026-08-11 | $0.250 | $0.750 |
Common questions
- How much does Qwen2.5 VL 72B Instruct cost per 1M tokens?
- Qwen2.5 VL 72B Instruct costs $0.800 per 1M input tokens and $1.00 per 1M output tokens on the standard tier. Cached input is $0.400 per 1M.
- What is Qwen2.5 VL 72B Instruct's context window?
- Qwen2.5 VL 72B Instruct accepts up to 128,000 tokens in a single request, and can return up to 115,200 output tokens.
- Is there a cheaper Qwen model than Qwen2.5 VL 72B Instruct?
- Yes. Qwen2.5 72B Instruct is $0.360 per 1M input tokens versus $0.800 for Qwen2.5 VL 72B Instruct, with a 32,768 token context window.
- Where does this Qwen2.5 VL 72B Instruct price come from?
- From the OpenRouter catalogue, because Alibaba (Qwen) does not publish machine-readable pricing. OpenRouter is a reseller, so its rate can differ from Alibaba (Qwen)'s own — verify before committing spend.
Head-to-head comparisons
Other Qwen models
- Qwen3.8 Max Prime $4.00
- Qwen3.8 Omni Flash $0.150
- Qwen3.8 Max (0902) $2.00
- Qwen3.8 Flash $0.150
- Qwen3.8 27B $0.420
- Qwen3.8 2.4T A95B $2.00
- Qwen3.7 Flash $0.030
- Qwen3.7 Plus $0.320
- Qwen3.7 Max $1.48
- Qwen3.5 Plus 2026-04-20 $0.300
- Qwen3.6 Flash $0.188
- Qwen3.6 35B A3B $0.150
- Qwen3.6 Max Preview $1.03
- Qwen3.6 27B $0.320
- Qwen3.6 Plus $0.325
- Qwen3.5-9B $0.100