Qwen/Qwen3-Next-80B-A3B-Instruct pricing
Over the past few months, we have observed increasingly clear trends toward scaling both total parameters and context lengths in the pursuit of more powerful and agentic artificial intelligence (AI). Qwen/Qwen3-Next-80B-A3B-Instruct from DeepInfra costs $0.090 per 1M input tokens and $1.10 per 1M output tokens. Context window 262,144 tokens.
API id Qwen/Qwen3-Next-80B-A3B-Instruct · DeepInfra · source · updated 2026-09-25 · general-purpose
$0.090
Input / 1M tokens
$1.10
Output / 1M tokens
—
Cached input / 1M
262K
Context window
Qwen/Qwen3-Next-80B-A3B-Instruct specifications
- Provider
- DeepInfra
- API id
Qwen/Qwen3-Next-80B-A3B-Instruct- Model type
- general-purpose
- Context window
- 262K tokens
- Max output
- not published
- Currency
- USD
- Price source
- Provider API
- Last checked
- 2026-09-25
Qwen3 Next 80B A3B Instruct across providers
The same model, priced by every seller we track. Ranked by the cost of 1M input + 1M output tokens, standard tier; savings measured against buying from the vendor directly.
| Provider | Input /1M | Cached /1M | Output /1M | 1M in + 1M out | vs vendor |
|---|---|---|---|---|---|
| $0.090 | — | $1.10 | $1.19 | −1% | |
Alibaba (Qwen)Vendor | $0.100 | $0.070 | $1.10 | $1.20 | — |
OpenRouterHost | $0.100 | $0.070 | $1.10 | $1.20 | — |
What Qwen/Qwen3-Next-80B-A3B-Instruct costs in practice
| Workload | Input tokens | Output tokens | Estimated cost |
|---|---|---|---|
| Short chat turn | 1,000 | 500 | $0.0006 |
| Document summary | 50,000 | 2,000 | $0.0067 |
| 1M tokens in, 200K out | 1,000,000 | 200,000 | $0.31 |
Estimates use the standard tier and ignore caching discounts, so real bills with repeated prompt prefixes are typically lower.
Common questions
- How much does Qwen/Qwen3-Next-80B-A3B-Instruct cost per 1M tokens?
- Qwen/Qwen3-Next-80B-A3B-Instruct costs $0.090 per 1M input tokens and $1.10 per 1M output tokens on the standard tier.
- What is Qwen/Qwen3-Next-80B-A3B-Instruct's context window?
- Qwen/Qwen3-Next-80B-A3B-Instruct accepts up to 262,144 tokens in a single request.
- Is there a cheaper DeepInfra model than Qwen/Qwen3-Next-80B-A3B-Instruct?
- Yes. mistralai/Mistral-Nemo-Instruct-2407 is $0.019 per 1M input tokens versus $0.090 for Qwen/Qwen3-Next-80B-A3B-Instruct, with a 131,072 token context window.
- Where does this Qwen/Qwen3-Next-80B-A3B-Instruct price come from?
- From the OpenRouter catalogue, because DeepInfra does not publish machine-readable pricing. OpenRouter is a reseller, so its rate can differ from DeepInfra's own — verify before committing spend.
Head-to-head comparisons
Other DeepInfra models
- XiaomiMiMo/MiMo-V2.6-Flash $0.140
- anthropic/claude-fable-5 $10.00
- Qwen/Qwen3-VL-235B-A22B-Instruct $0.200
- Qwen/Qwen3-14B $0.120
- Qwen/Qwen3-Max-Thinking $1.20
- ByteDance/Seed-2.0-mini $0.100
- deepseek-ai/DeepSeek-V4-Flash-0731 $0.060
- zai-org/GLM-5.3-Flash $0.150
- anthropic/claude-opus-5-5 $4.00
- meta-llama/Llama-4-Scout-17B-16E-Instruct $0.100
- ibm-granite/granite-4.2-8b $0.060
- Qwen/Qwen3-30B-A3B $0.120
- openai/gpt-oss-120b-Ultra $0.200
- ByteDance/Seed-2.0-pro $0.500
- thinkingmachines/Inkling $0.950
- zai-org/GLM-4.6 $0.500