/ Models

← Media models

Stable Audio 3 Small Music Text to Audio

Stable Audio 3 Small Music is a 459 million parameter latent diffusion model that generates full stereo music compositions up to 2 minutes from text prompts, lightweight enough for on-device deployment.

We have not generated an example for this one yet. Everything shown on this site is ours and hosted by us, so where there is no example there is simply nothing to show.

Takes
text
Makes
music
Release date
2026-06-03
Licence
commercial

What it costs

The supplier does not state a price on the model itself — check the link below.

Where it runs

Runs through fal, on one key that covers most of the open and licensed media models.

Licence: commercial

Open at the supplier →

← every media model