/ Models

Stable Audio 3 Medium Audio to Audio

Stable Audio 3 Medium audio-to-audio is a 1.4 billion parameter latent diffusion model that transforms an input audio clip into new stereo variations up to 6 minutes guided by a text prompt.

Takes
audio
Makes
audio-edit
Release date
2026-06-03
Licence
commercial

What it costs

The supplier does not state a price on the model itself — check the supplier's page.

Where it runs

Runs through fal, on one key that covers most of the open and licensed media models.

Licence: commercial

Open at the supplier →