Fresh paint, wet floors — we're rebuilding CostOfToken into a cross-provider price comparison. Some corners are still under construction; the full version is coming up soon.

Qwen3 VL 8B Instruct pricing

Qwen3-VL-8B-Instruct is a multimodal vision-language model from the Qwen3-VL series, built for high-fidelity understanding and reasoning across text, images, and video. Qwen3 VL 8B Instruct from Alibaba (Qwen) costs $0.117 per 1M input tokens and $0.455 per 1M output tokens. Context window 262,144 tokens.

API id qwen3-vl-8b-instruct · Alibaba (Qwen) · source · updated 2026-09-25 · general-purpose

$0.117
Input / 1M tokens
$0.455
Output / 1M tokens
—
Cached input / 1M
262K
Context window

Qwen3 VL 8B Instruct specifications

API id
qwen3-vl-8b-instruct
Model type
general-purpose
Context window
262K tokens
Max output
33K tokens
Currency
USD
Price source
Provider API
Last checked
2026-09-25

Tags

  • vision

Qwen3 VL 8B Instruct across providers

The same model, priced by every seller we track. Ranked by the cost of 1M input + 1M output tokens, standard tier; savings measured against buying from the vendor directly.

ProviderInput /1MCached /1MOutput /1M1M in + 1M outvs vendor
Alibaba (Qwen)VendorCheapestviewing
$0.117—$0.455$0.57—
$0.117—$0.455$0.57—

What Qwen3 VL 8B Instruct costs in practice

Estimated cost at common usage volumes
WorkloadInput tokensOutput tokensEstimated cost
Short chat turn1,000500$0.0003
Document summary50,0002,000$0.0068
1M tokens in, 200K out1,000,000200,000$0.21

Estimates use the standard tier and ignore caching discounts, so real bills with repeated prompt prefixes are typically lower.

Common questions

How much does Qwen3 VL 8B Instruct cost per 1M tokens?
Qwen3 VL 8B Instruct costs $0.117 per 1M input tokens and $0.455 per 1M output tokens on the standard tier.
What is Qwen3 VL 8B Instruct's context window?
Qwen3 VL 8B Instruct accepts up to 262,144 tokens in a single request, and can return up to 32,768 output tokens.
Is there a cheaper Qwen model than Qwen3 VL 8B Instruct?
Yes. Qwen2.5 7B Instruct is $0.100 per 1M input tokens versus $0.117 for Qwen3 VL 8B Instruct, with a 32,768 token context window.
Where does this Qwen3 VL 8B Instruct price come from?
From the OpenRouter catalogue, because Alibaba (Qwen) does not publish machine-readable pricing. OpenRouter is a reseller, so its rate can differ from Alibaba (Qwen)'s own — verify before committing spend.

Head-to-head comparisons

Other Qwen models

See all Qwen pricing →