Qwen3 VL 235B A22B Thinking pricing

Qwen3 VL 235B A22B Thinking from Alibaba (Qwen) costs $0.400 per 1M input tokens and $4.00 per 1M output tokens. Context window 131,072 tokens.

API id qwen3-vl-235b-a22b-thinking · Alibaba (Qwen) · source · updated 2026-08-12

$0.400
Input / 1M tokens
$4.00
Output / 1M tokens
Cached input / 1M
131K
Context window

What Qwen3 VL 235B A22B Thinking costs in practice

Estimated cost at common usage volumes
WorkloadInput tokensOutput tokensEstimated cost
Short chat turn1,000500$0.0024
Document summary50,0002,000$0.03
1M tokens in, 200K out1,000,000200,000$1.20

Estimates use the standard tier and ignore caching discounts, so real bills with repeated prompt prefixes are typically lower.

Common questions

How much does Qwen3 VL 235B A22B Thinking cost per 1M tokens?
Qwen3 VL 235B A22B Thinking costs $0.400 per 1M input tokens and $4.00 per 1M output tokens on the standard tier.
What is Qwen3 VL 235B A22B Thinking's context window?
Qwen3 VL 235B A22B Thinking accepts up to 131,072 tokens in a single request, and can return up to 32,768 output tokens.
Is there a cheaper Qwen model than Qwen3 VL 235B A22B Thinking?
Yes. Qwen2.5 72B Instruct is $0.360 per 1M input tokens versus $0.400 for Qwen3 VL 235B A22B Thinking, with a 32,768 token context window.
Where does this Qwen3 VL 235B A22B Thinking price come from?
From the OpenRouter catalogue, because Alibaba (Qwen) does not publish machine-readable pricing. OpenRouter is a reseller, so its rate can differ from Alibaba (Qwen)'s own — verify before committing spend.

Other Qwen models

See all Qwen pricing →