gpt-realtime-mini pricing
gpt-realtime-mini from OpenAI costs $0.600 per 1M input tokens and $2.40 per 1M output tokens. Cached input is $0.060 per 1M.
API id gpt-realtime-mini · OpenAI · source · updated 2026-09-25
We have not confirmed what kind of model this is. Its name suggests one thing and its pricing another, and we would rather say so than guess.
$0.600
Input / 1M tokens
$2.40
Output / 1M tokens
$0.060
Cached input / 1M
—
Context window
gpt-realtime-mini specifications
- Provider
- OpenAI
- API id
gpt-realtime-mini- Model type
- unknown
- Context window
- — tokens
- Max output
- not published
- Currency
- USD
- Price source
- First-party
- Last checked
- 2026-09-25
Tags
- audio
- fast
What gpt-realtime-mini costs in practice
| Workload | Input tokens | Output tokens | Estimated cost |
|---|---|---|---|
| Short chat turn | 1,000 | 500 | $0.0018 |
| Document summary | 50,000 | 2,000 | $0.03 |
| 1M tokens in, 200K out | 1,000,000 | 200,000 | $1.08 |
Estimates use the standard tier and ignore caching discounts, so real bills with repeated prompt prefixes are typically lower.
gpt-realtime-mini price history
| Recorded | Input /1M | Output /1M |
|---|---|---|
| 2026-08-25 | $0.600 | $2.40 |
| 2026-08-23 | $10.00 | — |
| 2026-08-22 | $0.600 | — |
Common questions
- How much does gpt-realtime-mini cost per 1M tokens?
- gpt-realtime-mini costs $0.600 per 1M input tokens and $2.40 per 1M output tokens on the standard tier. Cached input is $0.060 per 1M.
- What is gpt-realtime-mini's context window?
- OpenAI does not publish a context window for gpt-realtime-mini on its pricing page.
- Is there a cheaper OpenAI model than gpt-realtime-mini?
- Yes. omni-moderation-latest is free per 1M input tokens versus $0.600 for gpt-realtime-mini.
- Where does this gpt-realtime-mini price come from?
- Directly from OpenAI's published pricing page, re-read every day and normalized to USD per 1,000,000 tokens.