Model Details
Where is GPT Realtime a perfect fit?
Perfect Fit For:
- Voice-enabled assistants and interactive chatbots
- Real-time customer service or sales tools
- Multiplayer or collaborative AI experiences
- Live transcription, commentary, or tutoring applications
Quick Model Estimate
Your GPT Realtime Cost Estimate
π° Total Cost
β
for 1000 input + 1000 output tokens
Cost Breakdown
Prices are watched for changes and checked against the provider's own page.
Pricing
|
Provider
β
|
Modality
β
|
Service Tier
β
|
Input Price
(per 1M tokens) β |
Output Price
(per 1M tokens) β |
Cached Input
(per 1M tokens) β |
Context Size
β
|
View
|
|---|---|---|---|---|---|---|---|
OpenAI | Audio | Standard | $32.0000 | $64.0000 | $0.4000 | 32,000 tokens | β |
OpenAI | Image | Standard | $5.0000 | N/A | $0.5000 | 32,000 tokens | β |
OpenAI | Text | Standard | $4.0000 | $16.0000 | $0.4000 | 32,000 tokens | β |
Other Models in the GPT Realtime Family
FAQs about GPT Realtime
How much does GPT Realtime cost per 1M tokens?
GPT Realtime costs between $4.00 and $32.00 per million input tokens, and between $16.00 and $64.00 per million output tokens.
What does a typical workload cost with GPT Realtime?
1,000 requests of 2,000 input and 500 output tokens each β 2,000,000 input and 500,000 output tokens in total β costs $16.00 with GPT Realtime at its lowest rates. Use the calculator on this page for your own volumes.
Where can I use GPT Realtime?
GPT Realtime is available through OpenAI, from $4.00 per million input tokens.
What is GPT Realtime's context window?
GPT Realtime has a context window of 32,000 tokens, and returns up to 4,096 tokens in a single response.
Is GPT Realtime still available?
GPT Realtime is a legacy model and is scheduled for retirement on 20 January 2027. The replacement is GPT Realtime 2.1.
