Suno Premier's 30-minute upload versus Sume's audio limits
Suno Premier allows audio uploads up to 30 minutes. Sume's Timeline audio joins up to 1,800 s (30 min), while audio detach takes 1,800 s in and 900 s out.

Suno's pricing page lists audio uploads of up to 30 minutes on Premier ($24 a month, $19.20 a month billed annually). Sume's limits are different in kind: Timeline audio outputs up to 1,800 seconds (also 30 minutes), audio detach reads sources up to 1,800 seconds and outputs up to 900, and speech-to-text jobs run up to 10 minutes.
The limits side by side
| Limit | Value | Where |
|---|---|---|
| Suno Premier audio upload | 30 minutes | Suno pricing |
| Timeline audio output | 1,800 s | Sume docs |
| Audio detach source | 1,800 s | Sume docs |
| Audio detach output | 900 s | Sume docs |
| STT job | 600 s (10 min) | Sume STT |
Different jobs
Suno's upload feeds its own generation and Studio tools, which Premier unlocks with MIDI, effects and automation. Sume's endpoints are separate calls: pull audio from a video at $0.01, join or split at $0.01 per job, transcribe at $0.01 per audio minute.
- A 30-minute track at STT is three 10-minute jobs.
- Detach output capped at 900 s means a 30-minute source needs two ranges.
- Joined output may reach 1,800 s.
Splitting for STT
Timeline audio split takes 1 to 20 ranges, so a 30-minute file becomes three 600-second parts for $0.01, and transcribing 30 minutes is 30 x $0.01 = $0.30 in speech-to-text.
curl -X POST https://api.sume.com/v1/timeline-1.0/audio \
-H "Authorization: Bearer $SUME_API_KEY" \
-H "Content-Type: application/json" \
-H "Idempotency-Key: split-30min-001" \
-d '{
"operation": "split",
"url": "https://media.sume.com/artifacts/artf_demo/spine.wav",
"ranges": [{ "start": 0, "end": 600 }, { "start": 600, "end": 1200 }, { "start": 1200 }]
}'Where the fields are documented
The split call takes operation, one url and ranges[], each range being start and an optional end. See the Timeline audio docs.
Sources
Related posts
More in Comparisons
- Suno Speech beta: voice and music in one pass, or separate tracks?
Suno's Speech beta makes voice and music in one track. Its blog lists wandering accents and long pauses. When to prefer separate TTS, music and a timeline mix.
- Suno terms: commercial use needs an approved Pro or Premier download
Suno terms limit free and basic outputs to personal use; Pro and Premier may use them commercially via an approved download. What that means for a video ad.
- Suno v6 and label partners: what it means for an ad soundtrack
Suno's v6 post names WMG, BMG and Believe and upload screening. What a rights-minded team should check before using any AI music in a paid ad.
- Symphony Seedance 2.5 is invite-only: the API route today
TikTok Symphony with Seedance 2.5 is limited to select paid advertisers in select markets. The Sume API runs seedance-2.5 for 4 to 30 seconds today.
Written by Sume