AI Potluck
Back to Gap Map Model components / Image, video, 3D & music generation

Stable Audio

Stability AI
open weights / Overall score: 3.0

Stable Audio is Stability AI's family of diffusion models that generate music and sound effects from text. The Stable Audio 3 generation ships Small Music, Small SFX and Medium models that produce clips of up to two and six minutes and support editing and continuation, alongside the earlier Stable Audio Open models. The largest Stable Audio 3 model is not publicly downloadable; Stability offers it through its API and enterprise self-hosting.

The stable-audio-tools training and inference library is not part of this entry; it is a separate library in the registry.

Openness

3 high confidence
3.0
weights
open(the Stable Audio 3 Small and Medium checkpoints and Stable Audio Open download from Hugging Face, some behind an automatically granted click-through
data
described(trained on licensed and Creative Commons data, which is described and not released)
code
open(the training and inference pipeline is published in Stability-AI/stable-audio-3 under MIT)
license
Stability-AI-Community-License(every distributed checkpoint: free below USD 1M annual revenue, after which an enterprise license is required)

Every downloadable Stable Audio model is free to use, commercially included, for organizations under USD 1M in annual revenue, with a paid license above that. The training pipeline is published and the licensed training data is described but not released. Stability-AI-Community-License allows commercial use only within a bound, so the license tier is use_bounded.

Adoption

3 high confidence
3.0

Adoption is measured as Hugging Face downloads across the nine generation checkpoints, led by Stable Audio 3 Medium. The separate autoencoder repositories are left out.

Capability

3 medium confidence
3.0

Stable Audio 3 Large places in the upper half of the instrumental music arena, and the downloadable Medium model in its lower half. Suno and Lyria both place above the family's best model.

Verified 2026-09-26