Qwen3 VL 8B Instruct pricing

Qwen3 VL 8B Instruct from Alibaba (Qwen) costs $0.117 per 1M input tokens and $0.455 per 1M output tokens. Context window 262,144 tokens.

API id qwen3-vl-8b-instruct · Alibaba (Qwen) · source · updated 2026-08-11

$0.117
Input / 1M tokens
$0.455
Output / 1M tokens
Cached input / 1M
262K
Context window

What Qwen3 VL 8B Instruct costs in practice

Estimated cost at common usage volumes
WorkloadInput tokensOutput tokensEstimated cost
Short chat turn1,000500$0.0003
Document summary50,0002,000$0.0068
1M tokens in, 200K out1,000,000200,000$0.21

Estimates use the standard tier and ignore caching discounts, so real bills with repeated prompt prefixes are typically lower.

Common questions

How much does Qwen3 VL 8B Instruct cost per 1M tokens?
Qwen3 VL 8B Instruct costs $0.117 per 1M input tokens and $0.455 per 1M output tokens on the standard tier.
What is Qwen3 VL 8B Instruct's context window?
Qwen3 VL 8B Instruct accepts up to 262,144 tokens in a single request, and can return up to 32,768 output tokens.
Is there a cheaper Qwen model than Qwen3 VL 8B Instruct?
Yes. Qwen2.5 7B Instruct is $0.100 per 1M input tokens versus $0.117 for Qwen3 VL 8B Instruct, with a 32,768 token context window.
Where does this Qwen3 VL 8B Instruct price come from?
From the OpenRouter catalogue, because Alibaba (Qwen) does not publish machine-readable pricing. OpenRouter is a reseller, so its rate can differ from Alibaba (Qwen)'s own — verify before committing spend.

Other Qwen models

See all Qwen pricing →