Stable Audio 3 Small SFX is a 459 million parameter latent diffusion model that generates high-quality sound effects from text prompts, designed for on-device deployment on mobile phones and consumer laptops.
fal-ai/stable-audio-3/small/sfx/text-to-audio
Stable Audio 3 Small SFX is a 459 million parameter latent diffusion model that generates high-quality sound effects from text prompts, designed for on-device deployment on mobile phones and consumer laptops.
22-second cinematic sound-design scene: a small harbor dawn bed with buoy bells, rope strain, lapping water, and gulls far away. Keep it detailed, spatial, and non-musical.
{
"seed": 737778,
"audio": {
"url": "https://v3b.fal.media/files/b/0a9ba22c/dYTrhEYuzfaR5C40P1IWZ_tmp0xjj98n5.mp3",
"file_name": "tmp0xjj98n5.mp3",
"file_size": 529806,
"content_type": "application/octet-stream"
},
"prompt": "22-second cinematic sound-design scene: a small harbor dawn bed with buoy bells, rope strain, lapping water, and gulls far away. Keep it detailed, spatial, and non-musical."
}