Text In, Sound Effect Out
Send a plain-English description — “glass bottle shattering on a concrete floor” — and the API returns an audio file of that sound. The model is Stable Audio Open 1.0 from Stability AI, run on ModelsLab’s own GPUs with 80 diffusion steps at 44.1 kHz stereo. You choose the length (3 to 15 seconds) and the format (mp3, wav or flac).
Billing is per second of audio: $0.001 a second with a $0.0047 minimum per effect, drawn from your plan. On the $149/month Open Source Unlimited plan there is no per-effect charge at all, which makes it a flat price for teams that generate sound libraries in bulk.
ElevenLabs Sound Effects is on the same API key too, through /api/v7/voice/sound-generation with model_id eleven_sound_effect at $0.06 per generation. Request errors logged per endpoint are on the API status page.
- Stable Audio Open 1.0 over plain HTTP — no SDK to install
- $0.001 per second, $0.0047 minimum; a 10-second effect is $0.01
- 3 to 15 second clips in mp3, wav or flac, 44.1 kHz stereo
- About 13 seconds of GPU time per effect
- File URL in the response, fetch by id, or webhook delivery



