9to5Mac · Zac Hall ·

OpenAI launches three voice models in the API: GPT-Realtime-2 with GPT-5-class reasoning, GPT-Realtime-Whisper for transcription, and GPT-Realtime-Translate

OpenAI has just released three new realtime voice models that it says will "unlock a new class of voice apps for developers."

OpenAI launches three voice models in the API: GPT-Realtime-2 with GPT-5-class reasoning, GPT-Realtime-Whisper for transcription, and GPT-Realtime-Translate

Lead Source

More

OpenAI: OpenAI
The Next Web: The Next Web
TestingCatalog: TestingCatalog
Armin Ronacher's Thoughts …: Armin Ronacher's Thoughts …
Gizmodo: Gizmodo
Digit: Digit
The New Stack: The New Stack
Digital Trends: Digital Trends
Latent.Space: Latent.Space
MarkTechPost: MarkTechPost
The Rundown AI: The Rundown AI
The Asia Business Daily: The Asia Business Daily
TechCrunch: TechCrunch

Discussion

TechSnif Coverage

OpenAI Drops Three Real-Time Voice Models Into Its API

OpenAI unleashes a trio of voice models for developers: reasoning, transcription, and translation — all in real time.

OpenAI just shipped three new real-time voice models through its API, aiming to supercharge what developers can build with voice.

The lineup: GPT-Realtime-2 brings GPT-5-level reasoning to live voice interactions. GPT-Realtime-Whisper handles transcription. GPT-Realtime-Translate does exactly what the name suggests — real-time translation.

The company says the release will "unlock a new class of voice apps for developers." That's corporate-speak, but the underlying tech is legitimately broad. We're talking voice assistants, live translation tools, transcription engines, and whatever else devs dream up with GPT-5-class intelligence baked into the audio pipeline.

Three models. Three distinct jobs. All real-time. OpenAI is clearly betting that voice is the next major API battleground — and it just armed developers with significantly sharper tools to build on.