Fresh paint, wet floors — we're rebuilding CostOfToken into a cross-provider price comparison. Some corners are still under construction; the full version is coming up soon.

Qwen3.8 Omni Flash pricing

Qwen3.8 Omni Flash is an omni-modal reasoning model from Alibaba, the first Qwen model built around agentic capabilities with native audio-video understanding. Qwen3.8 Omni Flash from Alibaba (Qwen) costs $0.150 per 1M input tokens and $0.470 per 1M output tokens. Cached input is $0.016 per 1M. Context window 1,000,000 tokens.

API id qwen3.8-omni-flash · Alibaba (Qwen) · source · updated 2026-09-25 · general-purpose

$0.150
Input / 1M tokens
$0.470
Output / 1M tokens
$0.016
Cached input / 1M
1M
Context window

Qwen3.8 Omni Flash specifications

API id
qwen3.8-omni-flash
Model type
general-purpose
Context window
1M tokens
Max output
131K tokens
Currency
USD
Price source
Provider API
Last checked
2026-09-25

Tags

  • fast
  • vision

Qwen3.8 Omni Flash across providers

The same model, priced by every seller we track. Ranked by the cost of 1M input + 1M output tokens, standard tier; savings measured against buying from the vendor directly.

ProviderInput /1MCached /1MOutput /1M1M in + 1M outvs vendor
Alibaba (Qwen)VendorCheapestviewing
$0.150$0.016$0.470$0.62—
$0.150$0.016$0.470$0.62—

What Qwen3.8 Omni Flash costs in practice

Estimated cost at common usage volumes
WorkloadInput tokensOutput tokensEstimated cost
Short chat turn1,000500$0.0004
Document summary50,0002,000$0.0084
1M tokens in, 200K out1,000,000200,000$0.24

Estimates use the standard tier and ignore caching discounts, so real bills with repeated prompt prefixes are typically lower.

Common questions

How much does Qwen3.8 Omni Flash cost per 1M tokens?
Qwen3.8 Omni Flash costs $0.150 per 1M input tokens and $0.470 per 1M output tokens on the standard tier. Cached input is $0.016 per 1M.
What is Qwen3.8 Omni Flash's context window?
Qwen3.8 Omni Flash accepts up to 1,000,000 tokens in a single request, and can return up to 131,072 output tokens.
Is there a cheaper Qwen model than Qwen3.8 Omni Flash?
Yes. Qwen2.5 7B Instruct is $0.100 per 1M input tokens versus $0.150 for Qwen3.8 Omni Flash, with a 32,768 token context window.
Where does this Qwen3.8 Omni Flash price come from?
From the OpenRouter catalogue, because Alibaba (Qwen) does not publish machine-readable pricing. OpenRouter is a reseller, so its rate can differ from Alibaba (Qwen)'s own — verify before committing spend.

Head-to-head comparisons

Other Qwen models

See all Qwen pricing →