GPT Realtime Pricing - Cost Calculator

Track GPT-Realtime pricing across providers and set alerts for instant updates on OpenAI’s live, low-latency conversational AI model.

GPT Realtime is a legacy multimodal model from OpenAI, kept available since its release on 28 August 2025. It is due to be replaced by GPT Realtime 2.1 on 20 January 2027. A 32,000-token context window feeds responses of up to 4,096 tokens. It costs between $4.00 and $32.00 per million input tokens, depending on service tier.

Last updated Sep 27, 2026

Never miss an AI pricing or launch update

Get notified about new model launches, pricing changes, provider updates, deprecations, and cost-saving insights.

πŸ”’ We respect your privacy. Unsubscribe anytime.

Model Details

Released
Aug 28, 2025
Status
Legacysunset Jan 20, 2027
Context Length
32,000
Max Output
4,096
Superseded By

Where is GPT Realtime a perfect fit?

GPT-Realtime is designed for instant, bidirectional streaming conversations, enabling low-latency AI responses suitable for interactive and continuous dialogue experiences. It powers natural, voice-driven, or synchronous chat systems.

Perfect Fit For:
- Voice-enabled assistants and interactive chatbots
- Real-time customer service or sales tools
- Multiplayer or collaborative AI experiences
- Live transcription, commentary, or tutoring applications

Quick Model Estimate

(USD 32.0000 per 1M tokens)
(USD 64.0000 per 1M tokens)

Your GPT Realtime Cost Estimate

πŸ’° Total Cost

β€”

for 1000 input + 1000 output tokens

πŸ“₯ Input (1000 Γ— $32.000000) β€”
πŸ“€ Output (1000 Γ— $64.000000) β€”

Cost Breakdown

πŸ“₯ Input πŸ“€ Output

Prices are watched for changes and checked against the provider's own page.

Pricing

Provider ↕
Modality ↕
Service Tier ↕
Input Price
(per 1M tokens)
↕
Output Price
(per 1M tokens)
↕
Cached Input
(per 1M tokens)
↕
Context Size ↕
View
OpenAI LogoOpenAIAudioStandard$32.0000$64.0000$0.400032,000 tokens→
OpenAI LogoOpenAIImageStandard$5.0000N/A$0.500032,000 tokens→
OpenAI LogoOpenAITextStandard$4.0000$16.0000$0.400032,000 tokens→

FAQs about GPT Realtime

How much does GPT Realtime cost per 1M tokens?

GPT Realtime costs between $4.00 and $32.00 per million input tokens, and between $16.00 and $64.00 per million output tokens.

What does a typical workload cost with GPT Realtime?

1,000 requests of 2,000 input and 500 output tokens each β€” 2,000,000 input and 500,000 output tokens in total β€” costs $16.00 with GPT Realtime at its lowest rates. Use the calculator on this page for your own volumes.

Where can I use GPT Realtime?

GPT Realtime is available through OpenAI, from $4.00 per million input tokens.

What is GPT Realtime's context window?

GPT Realtime has a context window of 32,000 tokens, and returns up to 4,096 tokens in a single response.

Is GPT Realtime still available?

GPT Realtime is a legacy model and is scheduled for retirement on 20 January 2027. The replacement is GPT Realtime 2.1.