Music bed under narration: 'no vocals' and 'no spoken word' clauses
Sume's Music docs say to end the prompt with 'Instrumental, no vocals' and add 'no spoken word' only when a narrator will talk over the track.

For a bed that sits under narration, end the Sume music prompt with "Instrumental, no vocals." and add "no spoken word" only when the track will go under a voiceover. That is the guidance in the Music docs. Put every exclusion in the positive prompt, because Music 1.0 does not support a negative prompt.
Source: Music 1.0 docs (read 2026-10-06).
Why two clauses?
"No vocals" keeps a sung line out of the track. "No spoken word" keeps spoken-word samples out, which matters when a narrator will speak over the bed. The docs say to add the second clause only under narration, so a plain music prompt stays shorter.
Why not use negative_prompt?
A non-empty negative_prompt returns HTTP 400 with public_reason=negative_prompt_unsupported. Omit the field or send an empty string. Lyria does not support negative prompts, so write what you want and what to leave out in the same sentence.
| Need | What to write | Field |
|---|---|---|
| No singing | Instrumental, no vocals | prompt |
| No spoken samples under a voiceover | no spoken word | prompt |
| Avoid a style | Name it in the positive prompt | prompt |
| Negative prompt | Not supported; omit or send empty | negative_prompt |
What should I do next?
Generate the bed, then place it in a Timeline render under the voiceover with duck_db so it drops while the narrator speaks. A tone check by ear is still the final step.
Sources
More in Media tools
- Kling motion control job done: a signed Python webhook receiver
Submit a Sume Kling 3.0 motion-control job with mode webhook, then verify the HMAC signature in Python, reject an empty secret and answer fast. Runnable code.
- Loop a 30-second music bed under a 3-minute Reel: the body
A short bed can cover a 180-second Reel with soundtrack.loop, duck_db and a fade-out. One Timeline 1.0 body, the 3-minute bill, and the limits to respect.
- MAI-Voice-2.1 is not on Sume: make talking clips with Sume TTS
Microsoft MAI-Voice-2.1 launched 2026-10-01 but Sume does not list it. Here is the path that ships: Sume TTS audio, then H3 Max lip sync on a still.
- MiniMax H3 Max lip sync API: your first clip in Python
Submit a still and a Sume-hosted audio file to POST /v1/minimax/h3-max/lip-sync, poll the job, and read the result. Python stdlib, with the 5 to 14.8 s rule.
Written by Sume