fal-ai/stable-audio-3/medium/audio-to-audio

Stable Audio 3 Medium audio-to-audio is a 1.4 billion parameter latent diffusion model that transforms an input audio clip into new stereo variations up to 6 minutes guided by a text prompt.
Inference
Commercial use

Prompt examples

Examples are generated using the Stable Audio 3 Medium Audio to Audio. You can customize them by clicking on the "Playground" button.

Request 019e5f95-b91b-7fd2-ba44-f94aa0802022. Status 200. fal-ai/stable-audio-3/medium/audio-to-audio. 3mo ago
arcade funk slap bass sparkle
stable-audio-3/medium/audio-to-audio · 3mo ago
Request 019e5f70-64c0-77a1-8028-f2881b587614. Status 200. fal-ai/stable-audio-3/medium/audio-to-audio. 3mo ago
Transform the source into bright arcade funk instrumental with slap bass, talkbox-style synth lead, and tight disco claps; preserve the main rhythmic contour while changing instrumentation, space, and groove. No vocals.
stable-audio-3/medium/audio-to-audio · 3mo ago
Stable Audio 3 Medium (Audio to Audio) API on fal