Skip to content
EN
English 简体中文 soon 日本語 soon

Stable Audio 2

Stability AI's audio model

Visit official site

What Stable Audio is

Stable Audio is Stability AI's generative audio platform. It produces music and sound effects from text prompts or from audio you supply, and version 2 extended output to full tracks of up to three minutes with clear musical structure and stereo output at standard audio resolution. Audio-to-audio conversion and style reshaping are included, and an open weights release exists for local experimentation.

What you can do with it

  • Generate music and effects from prompts
  • Convert one piece of audio into another style
  • Produce longer structured tracks
  • Run the open model locally where licensing allows
  • Prototype sound design ideas

Who it is for

  • Music and film producers
  • Game sound designers
  • Researchers and local model users

What to watch out for

  • Licence terms differ between the hosted service and the open weights release; check both before commercial use
  • Training data provenance for generative audio is contested across the industry, so follow current legal developments
  • Longer tracks still need mastering for release
  • Local execution needs capable hardware

Pros & cons

✓ What we like

  • Longer structured output than earlier versions
  • Audio-to-audio conversion available
  • Open weights for local work

! What to watch out for

  • Licence differs per release channel
  • Industry-wide training data disputes
  • Mastering still required

FAQ

How long can tracks be?

Version 2 is described as producing full tracks up to three minutes.

Can I use audio as input?

Yes. Audio-to-audio conversion and style reshaping are described.

Is there a local option?

Stable Audio Open is mentioned for local experiments.

Last reviewed: 2026-09-18

More AI music creation tools

View all →

How we review