Fresh paint, wet floors — we're rebuilding CostOfToken into a cross-provider price comparison. Some corners are still under construction; the full version is coming up soon.

google/gemma-4-31B-it pricing

Gemma is a family of open models built by Google DeepMind. Gemma 4 models are multimodal, handling text and image input and generating text output. google/gemma-4-31B-it from DeepInfra costs $0.130 per 1M input tokens and $0.380 per 1M output tokens. Context window 262,144 tokens.

API id google/gemma-4-31B-it · DeepInfra · source · updated 2026-09-25 · general-purpose

$0.130
Input / 1M tokens
$0.380
Output / 1M tokens
—
Cached input / 1M
262K
Context window

google/gemma-4-31B-it specifications

Provider
DeepInfra
API id
google/gemma-4-31B-it
Model type
general-purpose
Context window
262K tokens
Max output
not published
Currency
USD
Price source
Provider API
Last checked
2026-09-25

Google: Gemma 4 31B across providers

The same model, priced by every seller we track. Ranked by the cost of 1M input + 1M output tokens, standard tier; savings measured against buying from the vendor directly.

ProviderInput /1MCached /1MOutput /1M1M in + 1M outvs vendor
OpenRouterHostCheapest
$0.090$0.050$0.340$0.43—
DeepInfraHostviewing
$0.130—$0.380$0.51—
Free routes$0

Free routes are rate-limited, can rotate off the roster without notice, and may differ in context, tools or caching from the paid door. Always confirm on the provider's page before depending on one.

What google/gemma-4-31B-it costs in practice

Estimated cost at common usage volumes
WorkloadInput tokensOutput tokensEstimated cost
Short chat turn1,000500$0.0003
Document summary50,0002,000$0.0073
1M tokens in, 200K out1,000,000200,000$0.21

Estimates use the standard tier and ignore caching discounts, so real bills with repeated prompt prefixes are typically lower.

Common questions

How much does google/gemma-4-31B-it cost per 1M tokens?
google/gemma-4-31B-it costs $0.130 per 1M input tokens and $0.380 per 1M output tokens on the standard tier.
What is google/gemma-4-31B-it's context window?
google/gemma-4-31B-it accepts up to 262,144 tokens in a single request.
Is there a cheaper DeepInfra model than google/gemma-4-31B-it?
Yes. deepseek-ai/DeepSeek-V4-Flash is $0.090 per 1M input tokens versus $0.130 for google/gemma-4-31B-it, with a 1,048,576 token context window.
Where does this google/gemma-4-31B-it price come from?
From the OpenRouter catalogue, because DeepInfra does not publish machine-readable pricing. OpenRouter is a reseller, so its rate can differ from DeepInfra's own — verify before committing spend.

Head-to-head comparisons

Other DeepInfra models

See all DeepInfra pricing →