google/gemini-2.5-flash pricing
Gemini 2.5 Flash is Google's latest thinking model, designed to tackle increasingly complex problems. It's capable of reasoning through their thoughts before responding, resulting in enhanced performance and improved accuracy. Gemini 2.5 Flash: best for balancing reasoning and speed. google/gemini-2.5-flash from DeepInfra costs $0.300 per 1M input tokens and $2.50 per 1M output tokens. Context window 1,000,000 tokens.
API id google/gemini-2.5-flash · DeepInfra · source · updated 2026-09-25 · general-purpose
google/gemini-2.5-flash specifications
- Provider
- DeepInfra
- API id
google/gemini-2.5-flash- Model type
- general-purpose
- Context window
- 1M tokens
- Max output
- not published
- Currency
- USD
- Price source
- Provider API
- Last checked
- 2026-09-25
Tags
- fast
Gemini 2.5 Flash across providers
The same model, priced by every seller we track. Ranked by the cost of 1M input + 1M output tokens, standard tier; savings measured against buying from the vendor directly.
| Provider | Input /1M | Cached /1M | Output /1M | 1M in + 1M out | vs vendor |
|---|---|---|---|---|---|
| $0.300 | $0.030 | $2.50 | $2.80 | — | |
| $0.300 | — | $2.50 | $2.80 | — | |
OpenRouterHost | $0.300 | $0.030 | $2.50 | $2.80 | — |
What google/gemini-2.5-flash costs in practice
| Workload | Input tokens | Output tokens | Estimated cost |
|---|---|---|---|
| Short chat turn | 1,000 | 500 | $0.0015 |
| Document summary | 50,000 | 2,000 | $0.02 |
| 1M tokens in, 200K out | 1,000,000 | 200,000 | $0.80 |
Estimates use the standard tier and ignore caching discounts, so real bills with repeated prompt prefixes are typically lower.
google/gemini-2.5-flash price history
| Recorded | Input /1M | Output /1M |
|---|---|---|
| 2026-08-27 | $0.300 | $2.50 |
| 2026-08-27 | $0.300 | $2.50 |
Common questions
- How much does google/gemini-2.5-flash cost per 1M tokens?
- google/gemini-2.5-flash costs $0.300 per 1M input tokens and $2.50 per 1M output tokens on the standard tier.
- What is google/gemini-2.5-flash's context window?
- google/gemini-2.5-flash accepts up to 1,000,000 tokens in a single request.
- Is there a cheaper DeepInfra model than google/gemini-2.5-flash?
- Yes. deepseek-ai/DeepSeek-V4-Flash is $0.090 per 1M input tokens versus $0.300 for google/gemini-2.5-flash, with a 1,048,576 token context window.
- Where does this google/gemini-2.5-flash price come from?
- From the OpenRouter catalogue, because DeepInfra does not publish machine-readable pricing. OpenRouter is a reseller, so its rate can differ from DeepInfra's own — verify before committing spend.
Head-to-head comparisons
Other DeepInfra models
- XiaomiMiMo/MiMo-V2.6-Flash $0.140
- anthropic/claude-fable-5 $10.00
- Qwen/Qwen3-VL-235B-A22B-Instruct $0.200
- Qwen/Qwen3-14B $0.120
- Qwen/Qwen3-Max-Thinking $1.20
- ByteDance/Seed-2.0-mini $0.100
- deepseek-ai/DeepSeek-V4-Flash-0731 $0.060
- zai-org/GLM-5.3-Flash $0.150
- anthropic/claude-opus-5-5 $4.00
- meta-llama/Llama-4-Scout-17B-16E-Instruct $0.100
- ibm-granite/granite-4.2-8b $0.060
- Qwen/Qwen3-30B-A3B $0.120
- openai/gpt-oss-120b-Ultra $0.200
- ByteDance/Seed-2.0-pro $0.500
- thinkingmachines/Inkling $0.950
- zai-org/GLM-4.6 $0.500