AI sound effects generator API: an open-weight SFX model, then Sume
Sume lists no sound-effects route. Stable Audio 3 Small SFX is open-weight, so you can run it yourself, import the files to Sume and mix them in Timeline 1.0.

Sume's docs list no sound-effects generator API. A model built for that job is open-weight: Stability AI says Stable Audio 3.0 Small SFX is on Hugging Face. If you want effect files today, run that model yourself, then bring the files into Sume and layer them in Timeline 1.0. Sume's own audio routes are music and speech.
The Stable Audio fact is from Stability's announcement, read 2026-09-29. The Sume facts are from the Music Router and Timeline 1.0 docs. We did not run Small SFX, so this post does not describe its output.
What does Sume list for audio?
| Need | Route |
|---|---|
| A music track | POST /v1/music-router/generate, ids sume/music-auto, lyria-3.5, lyria-3-pro |
| Mix a bed under a render | POST /v1/timeline-1.0/render with a soundtrack |
| A standalone effect such as a door slam | Not listed |
How do I get my own effect files into a Sume render?
Timeline 1.0 accepts only Sume-hosted media.sume.com files. Its docs say to import first with POST /v1/media-imports, and to send an Idempotency-Key on the render. Once imported, an effect file is a normal Sume audio file you can use as the render's soundtrack bed, which takes url, gain_db and loop.
Can I ask the music route for an effect?
You can write any description into a Music Router prompt, but the route makes a music track, priced at the fixed Music price per generation. It is not a substitute for a one-shot effect, and Sume's docs make no promise about isolated sounds. Listen before you rely on it.
What should I check before using an open-weight SFX model?
- Read the model's license on its own page. Stability names a Community License and an Enterprise License for organizations with $1M or more in revenue.
- Check the file format and length your render needs before importing.
- Keep the source model and prompt in your project notes, as you would for any generated asset.
What can the Timeline soundtrack bed do?
The soundtrack field takes a Sume-hosted url and optional gain_db, loop, fade_out_seconds (up to 10) and duck_db (0 to 20). Ducking lowers the bed under the main audio and needs a real audio spine, not silence. A short effect loops or fades like any other bed, so a single imported file can be placed once or repeated.
Timeline 1.0 charges $0.10 per output minute, rounded up, and its docs say no provider inference is involved, only worker ffmpeg.
What is the smallest workflow that works?
Generate or pick an effect file outside Sume, import it with POST /v1/media-imports, and reference its media.sume.com URL in a render's soundtrack. Run POST /v1/timeline-1.0/plan first: it is an unbilled compile preflight that checks the request before you pay for a render.
Sources
Related posts
More in Media tools
- Does Suno watermark songs? What you can and cannot confirm
Suno's v6 launch is on its blog, but no watermark detail was readable there. How to check Suno's terms, and what Sume's music files are as delivered.
- Sync music to the beat of a short video with AI
YouTube's Gemini editing can sync music to the beat. To do it yourself on Sume, read beat times from a clip's music track and place Timeline cuts on them.
- How to sync subtitles with video: offset, drift, re-time
Sync subtitles by shifting every timestamp when all lines are off by the same amount, or scaling them when the gap grows. How to fix both for good.
- Telegram video note limits: 1 minute, square, upload only
A Telegram video note is a rounded square MPEG4 up to 1 minute, and bots cannot send it by URL. What sendVideoNote takes and how to render a square clip.
Written by Sume