Switch language한국어
Back to the list

Stability AI Releases Stable Audio 3: A Family of Fast Latent Diffusion Models for Audio Generation and Editing

TL;DR AI

Key summary

2 min read
  1. Stability AI released open weights and a paper for Stable Audio 3, a family of latent diffusion audio models.

  2. The models generate 44.1 kHz stereo audio with variable lengths, text conditioning, and inpainting-based editing.

  3. Small and medium weights are available on Hugging Face, while the large model is enterprise-only.

  4. The release brings higher-quality open audio generation and editing, with smaller variants usable on consumer hardware.

Read the original