Switch language한국어
Back to the list

Introducing next-generation audio models in the API

TL;DR AI

Key summary

2 min read
  1. OpenAI has released next-generation audio models in its API.

  2. The new gpt-4o-transcribe and gpt-4o-mini-transcribe models improve speech recognition accuracy and language detection.

  3. The text-to-speech side also adds more natural, customizable speaking styles.

  4. OpenAI says the models outperform Whisper on benchmarks and are built for more reliable voice applications.

Read the original