GPT-Realtime-2.1 mini vs GPT-Realtime-2.1
Side-by-side comparison of API pricing, specs, benchmarks, and capabilities
| Specification |
by OpenAI
|
by OpenAI
|
|---|---|---|
| Overview | ||
| Status | Current mid tier | Current flagship |
| Released | Jul 6, 2026 | Jul 6, 2026 |
| Pricing per million tokens | ||
| Input | $0.600/Mtok | $4.00/Mtok |
| Output | $2.40/Mtok | $24.00/Mtok |
| Blended avg | $1.05/Mtok | $9.00/Mtok |
| Cached input | $0.060/Mtok | $0.400/Mtok |
| Price basis | OpenAI prices this model per 1M tokens on three separate modalities: text at $0.60 input / $0.06 cached input / $2.40 output, audio at $10.00 / $0.30 / $20.00, and image input at $0.80 / $0.08. The tracked figures are the TEXT rates, matching how the other realtime rows are tracked. | OpenAI prices this model per 1M tokens on three separate modalities: text at $4.00 input / $0.40 cached input / $24.00 output, audio at $32.00 / $0.40 / $64.00, and image input at $5.00 / $0.50. The tracked figures are the TEXT rates, which is how GPT-Realtime-2 is tracked too, so the two rows compare like for like. Every text and audio rate is identical to GPT-Realtime-2; only the model generation differs. |
| Specifications | ||
| Context window | 128K tokens | 128K tokens |
| Parameters | Proprietary | Proprietary |
| Speed (TPS) | — | — |
| Modalities | ||
| Input |
audiotextimage
|
audiotextimage
|
| Providers | ||
| Available from |
OpenAI — $0.600/$2.40/Mtok
|
OpenAI — $4.00/$24.00/Mtok
|
Cost at scale
1M tokens · 50/50 input/output| Volume | GPT-Realtime-2.1 mini | GPT-Realtime-2.1 | Savings |
|---|---|---|---|
| 1M tokens | $1.05 | $9 | $7.95 (88.3%) |
| 10M tokens | $10.5 | $90 | $79.5 (88.3%) |
| 100M tokens | $105 | $900 | $795 (88.3%) |
| 1000M tokens | $1050 | $9000 | $7950 (88.3%) |
When to pick which
distilled from the pricing and spec data above- cost dominates: $1.05/Mtok blended vs $9.00 — 88.3% less on the same 50/50 token mix
- this pairing gives GPT-Realtime-2.1 no edge on price, context, cache discount, or measured performance — pick it only for qualitative fit (ecosystem, compliance, model behavior on your prompts)
The cheaper option here — GPT-Realtime-2.1 mini costs $1.05/Mtok blended on OpenAI.
Summary
GPT-Realtime-2.1 mini by OpenAI costs $0.600/Mtok input and $2.40/Mtok output, with a 128K-token context window. It supports audio, text, image input.
GPT-Realtime-2.1 by OpenAI costs $4.00/Mtok input and $24.00/Mtok output, with a 128K-token context window. It supports audio, text, image input.
On a blended cost basis, GPT-Realtime-2.1 mini is 88.3% cheaper than GPT-Realtime-2.1.
Note: Pricing is per million tokens. Actual costs vary with usage patterns, prompt caching, and batch discounts. Always verify against official provider pricing pages.