/ Models

Stable Audio 3 Small Music Text to Audio

Stable Audio 3 Small Music is a 459 million parameter latent diffusion model that generates full stereo music compositions up to 2 minutes from text prompts, lightweight enough for on-device deployment.

Takes
text
Makes
music
Release date
2026-06-03
Licence
commercial

What it costs

The supplier does not state a price on the model itself — check the supplier's page.

Where it runs

Runs through fal, on one key that covers most of the open and licensed media models.

Licence: commercial

Open at the supplier →