Openai
0.00%Context—
Max output—
p50 latency—
Throughput—
Input price$0.0002
Output price$0
No training
openai/gpt-realtime-whisper
GPT Realtime Whisper is a streaming speech-to-text model for applications that need low-latency transcript deltas from live audio. It is designed for realtime use cases where developers need to tune latency and accuracy. GPT Realtime Whisper is priced by audio duration rather than text tokens.
USD per one million tokens, using current provider data.
Reported availability over the last 24 hours; not a contractual SLA.
Compare current pricing, performance, and privacy across every available provider.
| Provider | Maximum context | Max output | p50 latency | p50 throughput | 24h uptime | Input / 1M tokens | Output / 1M tokens | Privacy |
|---|---|---|---|---|---|---|---|---|
Openai | — | — | — | — | 0.000% | $0.0002 | $0 |
Available modalities and capabilities across inference providers.
openai/gpt-realtime-whisperMore options from the same creator or model category.