Introducing Stable Audio 3 and SAME (Semantically-Aligned Music Autoencoder)
·2026.05.20 23:43
Key point
Released Stable Audio 3, a fast latent diffusion model for variable-length audio generation and editing.
Details
Stable Audio 3 is a fast family of latent diffusion models that support variable-length audio generation and editing.
The model is available in three versions—small, medium, and large—and comes with the newly introduced SAME (Semantically-Aligned Music Autoencoder) technology.