A/B test two AI music beds on one ad: cost and the Timeline steps
Make two beds for one ad at $0.125 each, render both cuts on Timeline at $0.10 per minute, and compare. A cost table for a 30-second ad.

Testing two music beds on the same ad costs less than a coffee when the ad is 30 seconds long. Each bed is one generation at a fixed $0.125 on Sume's Music Router, and each rendered cut on Timeline 1.0 is $0.10 per output minute, rounded up, so a 30-second ad rounds to one minute: $0.10. Two beds and two renders come to $0.45.
Cost for one test
| Step | Count | Unit price | Subtotal |
|---|---|---|---|
| Music generation | 2 | $0.125 | $0.25 |
| Timeline render (30 s rounds up to 1 min) | 2 | $0.10 | $0.20 |
| Optional plan preflight | 2 | $0 (unbilled) | $0 |
| Total | $0.45 |
How to keep the test fair
Change one thing. Keep the video slots, voice and audio.gain_db identical and change only the soundtrack.url. Set the same gain_db and fade_out_seconds on both. If the ad has a voiceover, set the same duck_db on both too, because ducking changes how loud the bed feels more than the track does.
Write two briefs that differ on one axis, for example tempo or instrumentation, rather than two unrelated prompts. For Google's reference, Lyria 3.5 is listed at $0.08 per song on its pricing page, read on 2026-10-05; Sume's price is the figure you pay.
Running it
Call POST /v1/timeline-1.0/plan first for each cut: it is an unbilled preflight and catches errors such as soundtrack_fade_exceeds_output before you spend a render. Then post both renders, using a distinct Idempotency-Key per variant so a retry does not double charge. Label the two outputs A and B in your own storage and run them as separate ad sets; the platform's own split-test tool does the measuring. Budget a little extra for a third take if neither bed lands: each regeneration is another $0.125, and the 30-second render for the winning bed is still only $0.10. Because music is a single generation with no edit pass, a bed you almost like is replaced rather than adjusted, so keep the prompts saved next to each result.
Sources
Related posts
More in Use cases
- A4 poster at 300 dpi: 2480x3508 is over the GPT Image 2.5 pixel cap
A4 at 300 dpi is 2480x3508, 8,699,840 pixels, above the 8,294,400 cap. Render 2400x3392 on GPT Image 2.5, then scale up 3.4% in Pillow and tag it 300 dpi.
- ACCC Black Friday sweep: match the clip's end date to the sale end
ACCC Black Friday sweeps target limited-time claims and countdown timers that miss the real sale end. Burn the end date from one value. Read 2026-10-05.
- Add a logo to a product image: Ideogram 4.5 with two input references
Send the product photo first and the logo second in input_references on Sume, and say which is which in the prompt. Order matters: the first image is edited.
- Afrobeats background music for a fashion ad: rhythm brief and cost
Brief an instrumental Afrobeats bed for a fashion ad: tempo, percussion, bass, and a short loopable bar. $0.125 per take on Sume, mixed on Timeline for $0.10.
Written by Sume