Switch language한국어
Back to the list

OpenAI Releases Three Realtime Audio Models: GPT-Realtime-2, GPT-Realtime-Translate, and GPT-Realtime-Whisper in the Realtime API

TL;DR AI

Key summary

2 min read
  1. OpenAI launched three new Realtime API audio models: GPT-Realtime-2 for reasoning-heavy voice agents, GPT-Realtime-Translate for live speech translation, and GPT-Realtime-Whisper for transcription.

  2. The company also moved the Realtime API from beta to general availability, signaling production readiness.

  3. The update gives developers lower-friction tools to build real-time voice products with better latency control, reasoning, translation, and transcription.

Read the original