All Models

gpt-realtime-whisper

whisper

Streaming speech-to-text model for low-latency transcript deltas from live audio

Providers 1
Released May 7, 2026
Input Modalities audio
Output Modalities text

Available Providers (1)

Provider Model ID Input Cost Output Cost Context Max Output Docs
Vercel AI Gateway openai/gpt-realtime-whisper /MTok /MTok

Capabilities

Reasoning
Tool Calling
Attachments
Open Weights
Structured Output