Qwen3 VL 8B Instruct pricing
Qwen3 VL 8B Instruct from Alibaba (Qwen) costs $0.117 per 1M input tokens and $0.455 per 1M output tokens. Context window 262,144 tokens.
API id qwen3-vl-8b-instruct · Alibaba (Qwen) · source · updated 2026-08-11
$0.117
Input / 1M tokens
$0.455
Output / 1M tokens
—
Cached input / 1M
262K
Context window
What Qwen3 VL 8B Instruct costs in practice
| Workload | Input tokens | Output tokens | Estimated cost |
|---|---|---|---|
| Short chat turn | 1,000 | 500 | $0.0003 |
| Document summary | 50,000 | 2,000 | $0.0068 |
| 1M tokens in, 200K out | 1,000,000 | 200,000 | $0.21 |
Estimates use the standard tier and ignore caching discounts, so real bills with repeated prompt prefixes are typically lower.
Common questions
- How much does Qwen3 VL 8B Instruct cost per 1M tokens?
- Qwen3 VL 8B Instruct costs $0.117 per 1M input tokens and $0.455 per 1M output tokens on the standard tier.
- What is Qwen3 VL 8B Instruct's context window?
- Qwen3 VL 8B Instruct accepts up to 262,144 tokens in a single request, and can return up to 32,768 output tokens.
- Is there a cheaper Qwen model than Qwen3 VL 8B Instruct?
- Yes. Qwen2.5 7B Instruct is $0.100 per 1M input tokens versus $0.117 for Qwen3 VL 8B Instruct, with a 32,768 token context window.
- Where does this Qwen3 VL 8B Instruct price come from?
- From the OpenRouter catalogue, because Alibaba (Qwen) does not publish machine-readable pricing. OpenRouter is a reseller, so its rate can differ from Alibaba (Qwen)'s own — verify before committing spend.
Other Qwen models
- Qwen3.8 Max $2.00
- Qwen3.7 Flash $0.030
- Qwen3.7 Plus $0.320
- Qwen3.7 Max $1.48
- Qwen3.5 Plus 2026-04-20 $0.300
- Qwen3.6 Flash $0.188
- Qwen3.6 35B A3B $0.150
- Qwen3.6 Max Preview $1.03
- Qwen3.6 27B $0.600
- Qwen3.6 Plus $0.325
- Qwen3.5-9B $0.100
- Qwen3.5-35B-A3B $0.140
- Qwen3.5-27B $0.195
- Qwen3.5-122B-A10B $0.290
- Qwen3.5-Flash $0.065
- Qwen3.5 Plus 2026-02-15 $0.260