GPT-4o Realtime Preview Pricing - Cost Calculator

Track GPT-4o-Realtime-Preview pricing across providers and get alerts for cost updates on OpenAI’s experimental streaming AI model.

GPT-4o Realtime Preview, a multimodal model from OpenAI released 1 October 2024, has been deprecated. It was replaced by GPT Realtime 1.5 on 7 May 2026. The context window holds 128,000 tokens, with up to 16,384 returned per response. It costs between $5.00 and $40.00 per million input tokens, depending on service tier.

Last updated Sep 27, 2026

Never miss an AI pricing or launch update

Get notified about new model launches, pricing changes, provider updates, deprecations, and cost-saving insights.

πŸ”’ We respect your privacy. Unsubscribe anytime.

Model Details

Released
Oct 1, 2024
Status
Deprecated(on May 07, 2026)
Context Length
128,000
Max Output
16,384
Superseded By

Where is GPT-4o Realtime Preview a perfect fit?

GPT-4o-Realtime-Preview is an early access, streaming-enabled multimodal model built for interactive, real-time AI conversations. It delivers rapid, continuous responses across text, vision, and audio inputs, ideal for latency-sensitive use cases.

Perfect Fit For:
- Real-time conversational AI and voice assistants
- Interactive customer support and sales tools
- Live educational, translation, or co-creation experiences
- AI-driven meetings or commentary platforms

Quick Model Estimate

(USD 40.0000 per 1M tokens)
(USD 80.0000 per 1M tokens)

Your GPT-4o Realtime Preview Cost Estimate

πŸ’° Total Cost

β€”

for 1000 input + 1000 output tokens

πŸ“₯ Input (1000 Γ— $40.000000) β€”
πŸ“€ Output (1000 Γ— $80.000000) β€”

Cost Breakdown

πŸ“₯ Input πŸ“€ Output

Prices are watched for changes and checked against the provider's own page.

Pricing

Provider ↕
Modality ↕
Service Tier ↕
Input Price
(per 1M tokens)
↕
Output Price
(per 1M tokens)
↕
Cached Input
(per 1M tokens)
↕
Context Size ↕
View
OpenAI LogoOpenAIAudioStandard$40.0000$80.0000$2.5000128,000 tokens→
OpenAI LogoOpenAITextStandard$5.0000$20.0000$2.5000128,000 tokens→

FAQs about GPT-4o Realtime Preview

How much does GPT-4o Realtime Preview cost per 1M tokens?

GPT-4o Realtime Preview costs between $5.00 and $40.00 per million input tokens, and between $20.00 and $80.00 per million output tokens.

What does a typical workload cost with GPT-4o Realtime Preview?

1,000 requests of 2,000 input and 500 output tokens each β€” 2,000,000 input and 500,000 output tokens in total β€” costs $20.00 with GPT-4o Realtime Preview at its lowest rates. Use the calculator on this page for your own volumes.

Where can I use GPT-4o Realtime Preview?

GPT-4o Realtime Preview is available through OpenAI, from $5.00 per million input tokens.

What is GPT-4o Realtime Preview's context window?

GPT-4o Realtime Preview has a context window of 128,000 tokens, and returns up to 16,384 tokens in a single response.

Is GPT-4o Realtime Preview still available?

GPT-4o Realtime Preview is deprecated and was retired on 7 May 2026. The replacement is GPT Realtime 1.5.