gpt-realtime-2 pricing
gpt-realtime-2 from OpenAI costs $4.00 per 1M input tokens and $24.00 per 1M output tokens. Cached input is $0.400 per 1M.
API id gpt-realtime-2 · OpenAI · source · updated 2026-09-26
We have not confirmed what kind of model this is. Its name suggests one thing and its pricing another, and we would rather say so than guess.
$4.00
Input / 1M tokens
$24.00
Output / 1M tokens
$0.400
Cached input / 1M
—
Context window
gpt-realtime-2 specifications
- Provider
- OpenAI
- API id
gpt-realtime-2- Model type
- unknown
- Context window
- — tokens
- Max output
- not published
- Currency
- USD
- Price source
- First-party
- Last checked
- 2026-09-26
Tags
- audio
What gpt-realtime-2 costs in practice
| Workload | Input tokens | Output tokens | Estimated cost |
|---|---|---|---|
| Short chat turn | 1,000 | 500 | $0.02 |
| Document summary | 50,000 | 2,000 | $0.25 |
| 1M tokens in, 200K out | 1,000,000 | 200,000 | $8.80 |
Estimates use the standard tier and ignore caching discounts, so real bills with repeated prompt prefixes are typically lower.
gpt-realtime-2 price history
| Recorded | Input /1M | Output /1M |
|---|---|---|
| 2026-08-25 | $4.00 | $24.00 |
| 2026-08-23 | $32.00 | — |
| 2026-08-22 | $4.00 | — |
Common questions
- How much does gpt-realtime-2 cost per 1M tokens?
- gpt-realtime-2 costs $4.00 per 1M input tokens and $24.00 per 1M output tokens on the standard tier. Cached input is $0.400 per 1M.
- What is gpt-realtime-2's context window?
- OpenAI does not publish a context window for gpt-realtime-2 on its pricing page.
- Is there a cheaper OpenAI model than gpt-realtime-2?
- Yes. omni-moderation-latest is free per 1M input tokens versus $4.00 for gpt-realtime-2.
- Where does this gpt-realtime-2 price come from?
- Directly from OpenAI's published pricing page, re-read every day and normalized to USD per 1,000,000 tokens.