Switch language한국어
Back to the list

OpenAI launches two versions of its speech transcription AI, GPT Transcribe, with customized support for real-time and asynchronous use

TL;DR AI

Key summary

2 min read
  1. OpenAI has added two transcription models to its API: GPT-Live-Transcribe for low-latency live speech and GPT-Transcribe for batch and completed audio files.

  2. The models are built to better handle context, noise, and real-world speech, and OpenAI says they beat earlier Whisper-based and live speech models in benchmarks.

  3. The launch expands OpenAI’s speech AI stack for developers, making it easier to build accurate captions, transcription, and voice-driven product features.

Read the original