Why every AI background track sounds the same: a seven-axis brief

If your AI music all sounds alike, the prompt is probably generic. Sume's Lyria brief has seven axes: emotion, genre, BPM, key, instruments, arc and era.

5 min readSume
All posts

AI background tracks sound alike when the prompts are alike. Sume's Music docs ask for a scene-specific brief with seven axes: emotion, genre, tempo as a number, key and mode, two to four instruments with texture, an arc with one named moment, and an era or production note. A prompt that names only a mood leaves every other choice to the model's defaults.

The axes and the example come from the Music 1.0 docs (read 2026-10-06).

What are the seven axes?

The axes are creative directions, not guaranteed output values, so check the audio you get.

Seven axes of a music brief (read 2026-10-06)
AxisExample from the docs
Emotionhushed, slightly melancholic
Genre or lineageneo-soul, drill, bossa nova, synthwave
Tempo as a number72 BPM, 142 BPM half-time
Key and modeD minor, E phrygian
Instruments with textureRhodes through tape wow, 808 with long glide
Arc with one named momentbreakdown at 0:20, full return at 0:28
Era or production1998 production, dry and close

How do I make a batch contrast?

For scenes in one project that should differ, change the broad genre family, the tempo by at least 12 BPM, and the lead instrument. If you want one consistent score, keep those fixed. You can also pass a scene still as image_url.

What else do I add?

End with one clause such as "Instrumental, no vocals." Put exclusions in the positive prompt, because a non-empty negative_prompt returns HTTP 400. The prompt can run up to 5,000 characters.

Sources

More in Media tools

All Media tools posts

Written by Sume