GPT-4o Mini Realtime Preview Pricing - Cost Calculator

Track GPT-4o-Mini-Realtime-Preview pricing and set alerts for OpenAI’s lightweight, low-latency model optimized for real-time AI interactions.

GPT-4o Mini Realtime Preview, a multimodal model from OpenAI released 17 December 2024, has been deprecated. It was replaced by GPT Realtime Mini on 7 May 2026. The context window holds 128,000 tokens, with up to 16,384 returned per response. It costs between $0.60 and $10.00 per million input tokens, depending on service tier.

Last updated Sep 27, 2026

Never miss an AI pricing or launch update

Get notified about new model launches, pricing changes, provider updates, deprecations, and cost-saving insights.

πŸ”’ We respect your privacy. Unsubscribe anytime.

Model Details

Released
Dec 17, 2024
Status
Deprecated(on May 07, 2026)
Context Length
128,000
Max Output
16,384
Superseded By

Where is GPT-4o Mini Realtime Preview a perfect fit?

GPT-4o-Mini-Realtime-Preview is a compact, cost-efficient version of OpenAI’s real-time model, balancing performance and affordability for continuous, low-latency AI interactions. It supports multimodal input and rapid streaming output.

Perfect Fit For:
- Voice chatbots and personal assistants
- Real-time tutoring or translation tools
- Lightweight interactive applications with strict latency needs
- Cost-sensitive prototypes or small-scale deployments

Quick Model Estimate

(USD 10.0000 per 1M tokens)
(USD 20.0000 per 1M tokens)

Your GPT-4o Mini Realtime Preview Cost Estimate

πŸ’° Total Cost

β€”

for 1000 input + 1000 output tokens

πŸ“₯ Input (1000 Γ— $10.000000) β€”
πŸ“€ Output (1000 Γ— $20.000000) β€”

Cost Breakdown

πŸ“₯ Input πŸ“€ Output

Prices are watched for changes and checked against the provider's own page.

Pricing

Provider ↕
Modality ↕
Service Tier ↕
Input Price
(per 1M tokens)
↕
Output Price
(per 1M tokens)
↕
Cached Input
(per 1M tokens)
↕
Context Size ↕
View
OpenAI LogoOpenAIAudioStandard$10.0000$20.0000$0.3000128,000 tokens→
OpenAI LogoOpenAITextStandard$0.6000$2.4000$0.3000128,000 tokens→

FAQs about GPT-4o Mini Realtime Preview

How much does GPT-4o Mini Realtime Preview cost per 1M tokens?

GPT-4o Mini Realtime Preview costs between $0.60 and $10.00 per million input tokens, and between $2.40 and $20.00 per million output tokens.

What does a typical workload cost with GPT-4o Mini Realtime Preview?

1,000 requests of 2,000 input and 500 output tokens each β€” 2,000,000 input and 500,000 output tokens in total β€” costs $2.40 with GPT-4o Mini Realtime Preview at its lowest rates. Use the calculator on this page for your own volumes.

Where can I use GPT-4o Mini Realtime Preview?

GPT-4o Mini Realtime Preview is available through OpenAI, from $0.60 per million input tokens.

What is GPT-4o Mini Realtime Preview's context window?

GPT-4o Mini Realtime Preview has a context window of 128,000 tokens, and returns up to 16,384 tokens in a single response.

Is GPT-4o Mini Realtime Preview still available?

GPT-4o Mini Realtime Preview is deprecated and was retired on 7 May 2026. The replacement is GPT Realtime Mini.