Switch language한국어
Back to the list

Advancing voice intelligence with new models in the API

TL;DR AI

Key summary

2 min read
  1. OpenAI launched three new audio models in its API for voice applications.

  2. The lineup includes GPT-Realtime-2 for reasoning-focused real-time voice, GPT-Realtime-Translate for live speech translation, and GPT-Realtime-Whisper for streaming speech-to-text.

  3. These models give developers more tools to build voice agents that can listen, reason, translate, transcribe, and respond in real time.

  4. The launch expands what’s possible for more natural, multilingual voice interfaces and live conversational products.

Read the original