Fresh paint, wet floors — we're rebuilding CostOfToken into a cross-provider price comparison. Some corners are still under construction; the full version is coming up soon.

Qwen3 VL 8B Thinking pricing

Qwen3-VL-8B-Thinking is the reasoning-optimized variant of the Qwen3-VL-8B multimodal model, designed for advanced visual and textual reasoning across complex scenes, documents, and temporal sequences. Qwen3 VL 8B Thinking from Alibaba (Qwen) costs $0.180 per 1M input tokens and $2.10 per 1M output tokens. Context window 131,072 tokens.

API id qwen3-vl-8b-thinking · Alibaba (Qwen) · source · updated 2026-09-26 · general-purpose

$0.180
Input / 1M tokens
$2.10
Output / 1M tokens
—
Cached input / 1M
131K
Context window

Qwen3 VL 8B Thinking specifications

API id
qwen3-vl-8b-thinking
Model type
general-purpose
Context window
131K tokens
Max output
33K tokens
Currency
USD
Price source
Provider API
Last checked
2026-09-26

Tags

  • reasoning
  • vision

Qwen3 VL 8B Thinking across providers

The same model, priced by every seller we track. Ranked by the cost of 1M input + 1M output tokens, standard tier; savings measured against buying from the vendor directly.

ProviderInput /1MCached /1MOutput /1M1M in + 1M outvs vendor
Alibaba (Qwen)VendorCheapestviewing
$0.180—$2.10$2.28—
$0.180—$2.10$2.28—

What Qwen3 VL 8B Thinking costs in practice

Estimated cost at common usage volumes
WorkloadInput tokensOutput tokensEstimated cost
Short chat turn1,000500$0.0012
Document summary50,0002,000$0.01
1M tokens in, 200K out1,000,000200,000$0.60

Estimates use the standard tier and ignore caching discounts, so real bills with repeated prompt prefixes are typically lower.

Common questions

How much does Qwen3 VL 8B Thinking cost per 1M tokens?
Qwen3 VL 8B Thinking costs $0.180 per 1M input tokens and $2.10 per 1M output tokens on the standard tier.
What is Qwen3 VL 8B Thinking's context window?
Qwen3 VL 8B Thinking accepts up to 131,072 tokens in a single request, and can return up to 32,768 output tokens.
Is there a cheaper Qwen model than Qwen3 VL 8B Thinking?
Yes. Qwen2.5 7B Instruct is $0.100 per 1M input tokens versus $0.180 for Qwen3 VL 8B Thinking, with a 32,768 token context window.
Where does this Qwen3 VL 8B Thinking price come from?
From the OpenRouter catalogue, because Alibaba (Qwen) does not publish machine-readable pricing. OpenRouter is a reseller, so its rate can differ from Alibaba (Qwen)'s own — verify before committing spend.

Head-to-head comparisons

Other Qwen models

See all Qwen pricing →