Fresh paint, wet floors — we're rebuilding CostOfToken into a cross-provider price comparison. Some corners are still under construction; the full version is coming up soon.

google/gemini-2.5-flash pricing

Gemini 2.5 Flash is Google's latest thinking model, designed to tackle increasingly complex problems. It's capable of reasoning through their thoughts before responding, resulting in enhanced performance and improved accuracy. Gemini 2.5 Flash: best for balancing reasoning and speed. google/gemini-2.5-flash from DeepInfra costs $0.300 per 1M input tokens and $2.50 per 1M output tokens. Context window 1,000,000 tokens.

API id google/gemini-2.5-flash · DeepInfra · source · updated 2026-09-25 · general-purpose

$0.300
Input / 1M tokens
$2.50
Output / 1M tokens
—
Cached input / 1M
1M
Context window

google/gemini-2.5-flash specifications

Provider
DeepInfra
API id
google/gemini-2.5-flash
Model type
general-purpose
Context window
1M tokens
Max output
not published
Currency
USD
Price source
Provider API
Last checked
2026-09-25

Tags

  • fast

Gemini 2.5 Flash across providers

The same model, priced by every seller we track. Ranked by the cost of 1M input + 1M output tokens, standard tier; savings measured against buying from the vendor directly.

ProviderInput /1MCached /1MOutput /1M1M in + 1M outvs vendor
GoogleVendorCheapest
$0.300$0.030$2.50$2.80—
DeepInfraHostviewing
$0.300—$2.50$2.80—
$0.300$0.030$2.50$2.80—

What google/gemini-2.5-flash costs in practice

Estimated cost at common usage volumes
WorkloadInput tokensOutput tokensEstimated cost
Short chat turn1,000500$0.0015
Document summary50,0002,000$0.02
1M tokens in, 200K out1,000,000200,000$0.80

Estimates use the standard tier and ignore caching discounts, so real bills with repeated prompt prefixes are typically lower.

google/gemini-2.5-flash price history

RecordedInput /1MOutput /1M
2026-08-27$0.300$2.50
2026-08-27$0.300$2.50

Common questions

How much does google/gemini-2.5-flash cost per 1M tokens?
google/gemini-2.5-flash costs $0.300 per 1M input tokens and $2.50 per 1M output tokens on the standard tier.
What is google/gemini-2.5-flash's context window?
google/gemini-2.5-flash accepts up to 1,000,000 tokens in a single request.
Is there a cheaper DeepInfra model than google/gemini-2.5-flash?
Yes. deepseek-ai/DeepSeek-V4-Flash is $0.090 per 1M input tokens versus $0.300 for google/gemini-2.5-flash, with a 1,048,576 token context window.
Where does this google/gemini-2.5-flash price come from?
From the OpenRouter catalogue, because DeepInfra does not publish machine-readable pricing. OpenRouter is a reseller, so its rate can differ from DeepInfra's own — verify before committing spend.

Head-to-head comparisons

Other DeepInfra models

See all DeepInfra pricing →