Switch language한국어
Back to the list

OpenAI launches new realtime voice and translation AI models

TL;DR AI

Key summary

2 min read
  1. OpenAI expanded its Realtime API with three new audio models: GPT-Realtime-2, GPT-Realtime-Translate, and GPT-Realtime-Whisper.

  2. GPT-Realtime-2 improves spoken reasoning and handles larger context, while the translation model supports 70+ input languages and 13 output languages.

  3. GPT-Realtime-Whisper adds live speech-to-text, giving developers a new option for streaming transcription.

  4. All three models are now priced and available for testing and integration in the Playground and API.

Read the original