Google: Gemma 4 31B pricing
Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Google: Gemma 4 31B from OpenRouter costs $0.090 per 1M input tokens and $0.340 per 1M output tokens. Cached input is $0.050 per 1M. Context window 262,144 tokens.
API id google/gemma-4-31b-it · OpenRouter · source · updated 2026-09-25 · general-purpose
Google: Gemma 4 31B specifications
- Provider
- OpenRouter
- API id
google/gemma-4-31b-it- Model type
- general-purpose
- Context window
- 262K tokens
- Max output
- 16K tokens
- Currency
- USD
- Price source
- Provider API
- Last checked
- 2026-09-25
Tags
- vision
Google: Gemma 4 31B across providers
The same model, priced by every seller we track. Ranked by the cost of 1M input + 1M output tokens, standard tier; savings measured against buying from the vendor directly.
| Provider | Input /1M | Cached /1M | Output /1M | 1M in + 1M out | vs vendor |
|---|---|---|---|---|---|
| $0.090 | $0.050 | $0.340 | $0.43 | — | |
DeepInfraHost | $0.130 | — | $0.380 | $0.51 | — |
- OpenRouter
google/gemma-4-31b-it:free
Free routes are rate-limited, can rotate off the roster without notice, and may differ in context, tools or caching from the paid door. Always confirm on the provider's page before depending on one.
What Google: Gemma 4 31B costs in practice
| Workload | Input tokens | Output tokens | Estimated cost |
|---|---|---|---|
| Short chat turn | 1,000 | 500 | $0.0003 |
| Document summary | 50,000 | 2,000 | $0.0052 |
| 1M tokens in, 200K out | 1,000,000 | 200,000 | $0.16 |
Estimates use the standard tier and ignore caching discounts, so real bills with repeated prompt prefixes are typically lower.
Common questions
- How much does Google: Gemma 4 31B cost per 1M tokens?
- Google: Gemma 4 31B costs $0.090 per 1M input tokens and $0.340 per 1M output tokens on the standard tier. Cached input is $0.050 per 1M.
- What is Google: Gemma 4 31B's context window?
- Google: Gemma 4 31B accepts up to 262,144 tokens in a single request, and can return up to 16,384 output tokens.
- Is there a cheaper OpenRouter model than Google: Gemma 4 31B?
- Yes. MythoMax 13B is $0.080 per 1M input tokens versus $0.090 for Google: Gemma 4 31B, with a 8,192 token context window.
- Where does this Google: Gemma 4 31B price come from?
- From the OpenRouter catalogue, because OpenRouter does not publish machine-readable pricing. OpenRouter is a reseller, so its rate can differ from OpenRouter's own — verify before committing spend.
Head-to-head comparisons
Other OpenRouter models
- Fireworks: Ember-1 $3.00
- Z.ai: GLM 5.3 Prime $2.80
- Qwen: Qwen3.8 Max Prime $4.00
- Space Bunny Alpha Free
- AionLabs: Aion 3.5 Mini $0.700
- AionLabs: Aion 3.5 $3.00
- Upstage: Solar Mini 4 $0.050
- Cohere: Command A+ $0.300
- OpenAI: GPT-6 Luna Pro $0.100
- OpenAI: GPT-6 Luna $0.100
- OpenAI: GPT-6 Sol Pro $2.00
- OpenAI: GPT-6 Sol $2.00
- Anthropic: Claude Opus 5.5 $4.00
- Xiaomi: MiMo-V2.6-Pro-UltraSpeed $4.35
- Xiaomi: MiMo-V2.6-Flash $0.140
- Xiaomi: MiMo-V2.6-Pro $0.435