Qwen3 VL 30B A3B Thinking pricing
Qwen3-VL-30B-A3B-Thinking is a multimodal model that unifies strong text generation with visual understanding for images and videos. Its Thinking variant enhances reasoning in STEM, math, and complex tasks. Qwen3 VL 30B A3B Thinking from Alibaba (Qwen) costs $0.200 per 1M input tokens and $2.40 per 1M output tokens. Context window 262,144 tokens.
API id qwen3-vl-30b-a3b-thinking · Alibaba (Qwen) · source · updated 2026-09-26 · general-purpose
$0.200
Input / 1M tokens
$2.40
Output / 1M tokens
—
Cached input / 1M
262K
Context window
Qwen3 VL 30B A3B Thinking specifications
- Provider
- Alibaba (Qwen)
- API id
qwen3-vl-30b-a3b-thinking- Model type
- general-purpose
- Context window
- 262K tokens
- Max output
- 33K tokens
- Currency
- USD
- Price source
- Provider API
- Last checked
- 2026-09-26
Tags
- reasoning
- vision
Qwen3 VL 30B A3B Thinking across providers
The same model, priced by every seller we track. Ranked by the cost of 1M input + 1M output tokens, standard tier; savings measured against buying from the vendor directly.
| Provider | Input /1M | Cached /1M | Output /1M | 1M in + 1M out | vs vendor |
|---|---|---|---|---|---|
| $0.200 | — | $2.40 | $2.60 | — | |
OpenRouterHost | $0.200 | — | $2.40 | $2.60 | — |
What Qwen3 VL 30B A3B Thinking costs in practice
| Workload | Input tokens | Output tokens | Estimated cost |
|---|---|---|---|
| Short chat turn | 1,000 | 500 | $0.0014 |
| Document summary | 50,000 | 2,000 | $0.01 |
| 1M tokens in, 200K out | 1,000,000 | 200,000 | $0.68 |
Estimates use the standard tier and ignore caching discounts, so real bills with repeated prompt prefixes are typically lower.
Common questions
- How much does Qwen3 VL 30B A3B Thinking cost per 1M tokens?
- Qwen3 VL 30B A3B Thinking costs $0.200 per 1M input tokens and $2.40 per 1M output tokens on the standard tier.
- What is Qwen3 VL 30B A3B Thinking's context window?
- Qwen3 VL 30B A3B Thinking accepts up to 262,144 tokens in a single request, and can return up to 32,768 output tokens.
- Is there a cheaper Qwen model than Qwen3 VL 30B A3B Thinking?
- Yes. Qwen2.5 7B Instruct is $0.100 per 1M input tokens versus $0.200 for Qwen3 VL 30B A3B Thinking, with a 32,768 token context window.
- Where does this Qwen3 VL 30B A3B Thinking price come from?
- From the OpenRouter catalogue, because Alibaba (Qwen) does not publish machine-readable pricing. OpenRouter is a reseller, so its rate can differ from Alibaba (Qwen)'s own — verify before committing spend.
Head-to-head comparisons
Other Qwen models
- Qwen3.8 Max Prime $4.00
- Qwen3.8 Omni Flash $0.150
- Qwen3.8 Max (0902) $2.00
- Qwen3.8 Flash $0.150
- Qwen3.8 27B $0.420
- Qwen3.8 2.4T A95B $2.00
- Qwen3.7 Flash $0.030
- Qwen3.7 Plus $0.320
- Qwen3.7 Max $1.48
- Qwen3.5 Plus 2026-04-20 $0.300
- Qwen3.6 Flash $0.188
- Qwen3.6 35B A3B $0.150
- Qwen3.6 Max Preview $1.03
- Qwen3.6 27B $0.320
- Qwen3.6 Plus $0.325
- Qwen3.5-9B $0.100