OpenAI adds three realtime voice models to API

Get the Tech newsletter
Daily tech — startups, AI labs, chips, the launches that shape the next decade. Free.
- OpenAI released three realtime voice models to its API, expanding its audio capabilities.
- GPT-Realtime-2 offers GPT‑5‑class reasoning, enabling complex conversational tasks.
- GPT-Realtime-Whisper provides real‑time transcription, turning spoken input into text.
- GPT-Realtime-Translate delivers live translation, allowing cross‑language voice interactions.
- OpenAI says the new suite will unlock a new class of voice apps for developers.
Why it matters: Developers gain three new realtime voice APIs, letting them build voice‑first apps and potentially capture new market revenue, while OpenAI redirects resources from Prism and Sora to this focused offering.
Ask SkimNews




