gpt-realtime-whisper Pricing - Cost Calculator

Track gpt-realtime-whisper pricing with real-time alerts, compare providers, and monitor cost changes across streaming speech transcription and realtime audio AI workflows.

Released 7 May 2026, gpt-realtime-whisper is a transcription model from OpenAI. It is available from OpenAI.

Last updated Sep 2, 2026

Never miss an AI pricing or launch update

Get notified about new model launches, pricing changes, provider updates, deprecations, and cost-saving insights.

๐Ÿ”’ We respect your privacy. Unsubscribe anytime.

Model Details

Released
May 7, 2026
Context Length
16,000
Max Output
2,000

Where is gpt-realtime-whisper a perfect fit?

GPT-Realtime-Whisper is OpenAI's dedicated streaming speech-to-text model, designed to return low-latency transcript updates while someone is still speaking. Developers can tune the trade-off between transcription latency and accuracy.
- Live captions and subtitles for calls, meetings, and broadcasts
- Realtime transcription while a user is actively speaking
- Voice assistants that need text transcripts before downstream processing
- Contact-center and customer-support call transcription
- Meeting notes and live conversation logging
- Accessibility applications providing immediate speech-to-text
- Applications requiring adjustable latency versus transcription accuracy
- Streaming audio workflows where waiting for a complete recording is impractical

Quick Model Estimate

(USD N/A per 1M tokens)
(USD 0.0170 per 1M tokens)

Your GPT-5 Cost Estimate

๐Ÿ’ฐ Total Cost

USD 3.00

for 1000 input + 1000 output tokens

๐Ÿ“ฅ Input (1000 ร— $N/A) USD 1.5000
๐Ÿ“ค Output (1000 ร— $0.017000) USD 1.5000

Cost Breakdown

๐Ÿ“ฅ Input 50% ๐Ÿ“ค Output 50%

Prices updated daily from official provider data.

Pricing

Provider โ†•
Modality โ†•
Service tier โ†•
Input Price
(per 1M tokens)
โ†•
Output Price
(per 1M tokens)
โ†•
Context Size โ†•
View
OpenAI LogoOpenAIAudioStandardN/A$0.017016,000 tokensโ†’

FAQs about gpt-realtime-whisper

What is gpt-realtime-whisper's context window?

gpt-realtime-whisper has a context window of 16,000 tokens, and returns up to 2,000 tokens in a single response.