Lyria 3.5 is single-turn and varies per call: keep the artifact
Google says Lyria generation is single-turn, not iteratively editable and varies between calls. Why to save every Sume Music artifact you like.

Google's Gemini API music page says music generation is a single-turn process, that iterative editing is not supported, and that results may vary between calls even with the same prompt. For Sume Music, which resolves sume/music-auto to Lyria 3.5 today, that means a track you like cannot be reproduced by re-sending the prompt. Download and store the artifact the first time.
What Google's page says
Model ids and the 30-second versus full-length split are from the same page.
| Item | What the page says |
|---|---|
| Models | lyria-3-clip-preview for 30-second clips; lyria-3.5 for full-length tracks |
| Editing | Single-turn; iterative editing not supported |
| Repeatability | Results may vary between calls, even with the same prompt |
| Output | 44.1 kHz stereo MP3; WAV on 3.5 |
| Watermark | SynthID |
What this means on Sume
Each Sume generation is a separate paid job at $0.125 per accepted generation, and job.request.routed_model records which engine ran. The result is an audio artifact on media.sume.com plus result.lyrics when present.
Re-rolling is therefore a deliberate cost. Generate a few variants in one batch with different Idempotency-Key values, pick one, and keep its job id.
import os, requests
for n in range(3):
r = requests.post(
"https://api.sume.com/v1/music-router/generate",
headers={"Authorization": f"Bearer {os.environ['SUME_API_KEY']}",
"Idempotency-Key": f"variant-{n}"},
json={"prompt": "A 30-second upbeat marimba jingle. Instrumental."},
)
print(n, r.status_code, r.json().get("id"))Editing without re-rolling
Trimming and joining do not need the model: use timeline audio to split a good section out or concatenate two takes.
Worked example
Two habits keep you out of trouble.
- Generate three variants for a key spot ($0.375 at $0.125 each) and choose one.
- Download the winning file and note its job id and
routed_model. - Fix small problems with timeline audio split or concat instead of re-rolling.
Checklist before you commit
Because Google says results vary between calls, do not use the same prompt as a stand-in for a saved file.
- Store the artifact in your own bucket.
- Write down the prompt.
- Re-check the Google page; the preview model ids may change.
Sources
Related posts
More in Models
- Lyria 3.5 in another language: prompt in it, then check the song
Google says Lyria 3.5 makes music in other languages when you prompt in that language. On Sume, write the brief and lyrics in it, then check the result.
- MAI-Image-2.6 output cap is 2,359,296 pixels; Sume sizes differ
MAI-Image-2.6 sets a 2,359,296-pixel ceiling and a 768-pixel minimum edge. Sume sets size per model with tiers, ratios and, for GPT models, custom pixels.
- MAI-Image-2.6 edits take 5 references; Sume takes 10 or 16
MAI-Image-2.6 in Foundry accepts up to five JPEG or PNG reference images per edit. On Sume, input_references tops out at 10, or 16 on GPT Image 2.5.
- MAI-Image-2.6 web_grounding flag: what Sume has instead
MAI-Image-2.6 can pull Bing results into an image when web_grounding is on. Sume has no such flag; here is how to pass current facts in the prompt instead.
Written by Sume