Qwen3 VL 8B Thinking pricing

Qwen3 VL 8B Thinking from Alibaba (Qwen) costs $0.180 per 1M input tokens and $2.10 per 1M output tokens. Context window 131,072 tokens.

API id qwen3-vl-8b-thinking · Alibaba (Qwen) · source · updated 2026-08-12

$0.180
Input / 1M tokens
$2.10
Output / 1M tokens
Cached input / 1M
131K
Context window

What Qwen3 VL 8B Thinking costs in practice

Estimated cost at common usage volumes
WorkloadInput tokensOutput tokensEstimated cost
Short chat turn1,000500$0.0012
Document summary50,0002,000$0.01
1M tokens in, 200K out1,000,000200,000$0.60

Estimates use the standard tier and ignore caching discounts, so real bills with repeated prompt prefixes are typically lower.

Common questions

How much does Qwen3 VL 8B Thinking cost per 1M tokens?
Qwen3 VL 8B Thinking costs $0.180 per 1M input tokens and $2.10 per 1M output tokens on the standard tier.
What is Qwen3 VL 8B Thinking's context window?
Qwen3 VL 8B Thinking accepts up to 131,072 tokens in a single request, and can return up to 32,768 output tokens.
Is there a cheaper Qwen model than Qwen3 VL 8B Thinking?
Yes. Qwen2.5 7B Instruct is $0.100 per 1M input tokens versus $0.180 for Qwen3 VL 8B Thinking, with a 32,768 token context window.
Where does this Qwen3 VL 8B Thinking price come from?
From the OpenRouter catalogue, because Alibaba (Qwen) does not publish machine-readable pricing. OpenRouter is a reseller, so its rate can differ from Alibaba (Qwen)'s own — verify before committing spend.

Other Qwen models

See all Qwen pricing →