OpenAI launches new realtime voice and translation AI models

TL;DR AI
2 min readKey summary
OpenAI expanded its Realtime API with three new audio models: GPT-Realtime-2, GPT-Realtime-Translate, and GPT-Realtime-Whisper.
GPT-Realtime-2 improves spoken reasoning and handles larger context, while the translation model supports 70+ input languages and 13 output languages.
GPT-Realtime-Whisper adds live speech-to-text, giving developers a new option for streaming transcription.
All three models are now priced and available for testing and integration in the Playground and API.


