Stable Audio 3 Medium Audio to Audio
Stable Audio 3 Medium audio-to-audio is a 1.4 billion parameter latent diffusion model that transforms an input audio clip into new stereo variations up to 6 minutes guided by a text prompt.
We have not generated an example for this one yet. Everything shown on this site is ours and hosted by us, so where there is no example there is simply nothing to show.
What it costs
The supplier does not state a price on the model itself — check the link below.
Where it runs
Runs through fal, on one key that covers most of the open and licensed media models.
Licence: commercial