Lyria 3.5 blocks artist-voice prompts: how to write briefs that pass
Google's Lyria 3.5 docs note that prompts asking for specific artist voices are blocked. Describe the sound instead, then run it through the Sume Music Router.

Describe the sound, not the singer. The Gemini API changelog for September 3 says Lyria 3.5 is generally available with full-length songs, natural vocals and fine control over duration and structure, and its docs note that prompts requesting specific artist voices are blocked.
What the changelog says
| Item | Stated detail |
|---|---|
| Status | Generally available (September 3) |
| Length | Full-length songs |
| Vocals | Natural vocals |
| Control | Fine duration and structure control |
| Restriction | Prompts asking for specific artist voices are blocked |
Rewrite a blocked prompt
Replace a name with the qualities you hear in it: register, texture, pace, mood, and language. The aim is a brief a human producer could follow without knowing who you had in mind.
| Instead of | Write |
|---|---|
| A named singer | Warm, breathy mezzo-soprano, close-mic, soft consonants |
| A named band sound | Four-piece indie rock, jangly guitars, steady mid-tempo drums |
| A named producer | Sparse beat, dry snare, deep sub bass, lots of space |
Steering structure on Sume
The Music Router takes a prompt of 1 to 5000 characters. Put exclusions in the positive prompt, because a non-empty negative_prompt is unsupported, and steer length and sections in the text, for example a 2-minute track or [0:00-0:30] Intro: …. duration and duration_seconds are rejected. With sume/music-auto the engine is Lyria 3.5 today, and job.request.routed_model tells you which one ran.
curl -X POST https://api.sume.com/v1/music-router/generate \
-H "Authorization: Bearer $SUME_API_KEY" \
-H "Content-Type: application/json" \
-H "Idempotency-Key: song-001" \
-d '{
"model": "sume/music-auto",
"prompt": "A 2-minute track. [0:00-0:30] Intro: soft piano. Warm, breathy mezzo-soprano vocal, English lyrics about a long drive home. Instrumental break at 1:00."
}'If a prompt is refused
Remove every proper name of a performer, then run again. Keep the genre, tempo, instruments and mood. Read result.lyrics on the finished job, which carries the model-reported lyrics or section map when present, to see how the structure was interpreted.
Sources
Related posts
More in Models
- MiniMax H3 limits: 9 images, 3 videos, 3 audio, file caps
MiniMax's H3 guide caps prompts at 7,000 characters and references at 9 images, 3 videos and 3 audio files. Cheat sheet with the Sume limits beside it.
- Nano Banana reference image limits: Lite, 2 and Pro compared
Nano Banana 2 Lite takes up to 14 object images, Nano Banana 2 takes 10 object, 4 character and 3 style, Pro takes 6 object and 5 character.
- PixVerse V6 native audio and camera work: what to check in an API
PixVerse's blog lists V6 with camera work and native audio, plus R2 and a $439M Series C total. How to test those claims against any video API's catalog.
- PixVerse V6 adds native audio: which Sume video ids make sound
PixVerse V6 ships camera work and native audio. The Sume video docs name no PixVerse id, so here is how to find models that generate audio and read the flag.
Written by Sume