Model Details
Where is gpt-realtime-whisper a perfect fit?
- Live captions and subtitles for calls, meetings, and broadcasts
- Realtime transcription while a user is actively speaking
- Voice assistants that need text transcripts before downstream processing
- Contact-center and customer-support call transcription
- Meeting notes and live conversation logging
- Accessibility applications providing immediate speech-to-text
- Applications requiring adjustable latency versus transcription accuracy
- Streaming audio workflows where waiting for a complete recording is impractical
Quick Model Estimate
Your GPT-5 Cost Estimate
๐ฐ Total Cost
USD 3.00
for 1000 input + 1000 output tokens
Cost Breakdown
Prices updated daily from official provider data.
Pricing
|
Provider
โ
|
Modality
โ
|
Service tier
โ
|
Input Price
(per 1M tokens) โ |
Output Price
(per 1M tokens) โ |
Context Size
โ
|
View
|
|---|---|---|---|---|---|---|
OpenAI | Audio | Standard | N/A | $0.0170 | 16,000 tokens | โ |
Other Models in the gpt-realtime Family
FAQs about gpt-realtime-whisper
What is gpt-realtime-whisper's context window?
gpt-realtime-whisper has a context window of 16,000 tokens, and returns up to 2,000 tokens in a single response.
