google/gemma-4-26B-A4B-it pricing
Efficient, MoE variant of Gemma 4. Gemma is a family of open models built by Google DeepMind. Gemma 4 models are multimodal, handling text and image input and generating text output. google/gemma-4-26B-A4B-it from DeepInfra costs $0.070 per 1M input tokens and $0.340 per 1M output tokens. Context window 262,144 tokens.
API id google/gemma-4-26B-A4B-it · DeepInfra · source · updated 2026-09-25 · general-purpose
google/gemma-4-26B-A4B-it specifications
- Provider
- DeepInfra
- API id
google/gemma-4-26B-A4B-it- Model type
- general-purpose
- Context window
- 262K tokens
- Max output
- not published
- Currency
- USD
- Price source
- Provider API
- Last checked
- 2026-09-25
Google: Gemma 4 26B A4B across providers
The same model, priced by every seller we track. Ranked by the cost of 1M input + 1M output tokens, standard tier; savings measured against buying from the vendor directly.
| Provider | Input /1M | Cached /1M | Output /1M | 1M in + 1M out | vs vendor |
|---|---|---|---|---|---|
| $0.090 | $0.050 | $0.300 | $0.39 | — | |
| $0.070 | — | $0.340 | $0.41 | — |
- OpenRouter
google/gemma-4-26b-a4b-it:free
Free routes are rate-limited, can rotate off the roster without notice, and may differ in context, tools or caching from the paid door. Always confirm on the provider's page before depending on one.
What google/gemma-4-26B-A4B-it costs in practice
| Workload | Input tokens | Output tokens | Estimated cost |
|---|---|---|---|
| Short chat turn | 1,000 | 500 | $0.0002 |
| Document summary | 50,000 | 2,000 | $0.0042 |
| 1M tokens in, 200K out | 1,000,000 | 200,000 | $0.14 |
Estimates use the standard tier and ignore caching discounts, so real bills with repeated prompt prefixes are typically lower.
Common questions
- How much does google/gemma-4-26B-A4B-it cost per 1M tokens?
- google/gemma-4-26B-A4B-it costs $0.070 per 1M input tokens and $0.340 per 1M output tokens on the standard tier.
- What is google/gemma-4-26B-A4B-it's context window?
- google/gemma-4-26B-A4B-it accepts up to 262,144 tokens in a single request.
- Is there a cheaper DeepInfra model than google/gemma-4-26B-A4B-it?
- Yes. mistralai/Mistral-Nemo-Instruct-2407 is $0.019 per 1M input tokens versus $0.070 for google/gemma-4-26B-A4B-it, with a 131,072 token context window.
- Where does this google/gemma-4-26B-A4B-it price come from?
- From the OpenRouter catalogue, because DeepInfra does not publish machine-readable pricing. OpenRouter is a reseller, so its rate can differ from DeepInfra's own — verify before committing spend.
Head-to-head comparisons
Other DeepInfra models
- XiaomiMiMo/MiMo-V2.6-Flash $0.140
- anthropic/claude-fable-5 $10.00
- Qwen/Qwen3-VL-235B-A22B-Instruct $0.200
- Qwen/Qwen3-14B $0.120
- Qwen/Qwen3-Max-Thinking $1.20
- ByteDance/Seed-2.0-mini $0.100
- deepseek-ai/DeepSeek-V4-Flash-0731 $0.060
- zai-org/GLM-5.3-Flash $0.150
- anthropic/claude-opus-5-5 $4.00
- meta-llama/Llama-4-Scout-17B-16E-Instruct $0.100
- ibm-granite/granite-4.2-8b $0.060
- Qwen/Qwen3-30B-A3B $0.120
- openai/gpt-oss-120b-Ultra $0.200
- ByteDance/Seed-2.0-pro $0.500
- thinkingmachines/Inkling $0.950
- zai-org/GLM-4.6 $0.500