Switch language한국어
Back to the list

OpenAI has new voice models that reason, translate, and transcribe as you speak

TL;DR AI

Key summary

2 min read
  1. OpenAI launched three new realtime voice models for reasoning, live translation, and streaming transcription.

  2. GPT-Realtime-2 handles live conversation reasoning, while GPT-Realtime-Translate converts speech from 70+ languages into 13 languages.

  3. GPT-Realtime-Whisper provides low-latency streaming transcription as speech happens.

  4. The models are available through the Realtime API with usage-based pricing and can be tested in Playground.

  5. The release gives developers new tools to build faster voice apps with reasoning, translation, and speech-to-text.

Read the original