Sume accepted job capacity by plan: 6, 24, 48, 120 and queue_full
Accepted capacity is concurrency plus queue: 6 on Free, 24 on Pro, 48 on Startup and 120 on Scale. Past it, submits get 429 queue_full until a job finishes.

How many jobs Sume accepts at once is the plan's concurrency plus its queue: 6 on Free, 24 on Pro, 48 on Startup and 120 on Scale and Enterprise. Another submit past that gets 429 queue_full, and it is safe to retry once a job completes.
This is a different limit from requests per minute, and it is the one that bites a batch.
Per plan
Numbers are from the generation admission page (read 2026-10-03).
| Plan | Concurrent | Queue | Accepted |
|---|---|---|---|
| Free | 1 | 5 | 6 |
| Pro | 4 | 20 | 24 |
| Startup | 8 | 40 | 48 |
| Scale | 20 | 100 | 120 |
| Enterprise | 20 | 100 | 120 |
Submit in waves
GET /v1/generation/admission-preview returns generation_limits, including wave_size_hint, which is the larger of 1 and the floor of 75 percent of the remaining queue capacity. It is a hint, not a promise. The in-flight budget is concurrency minus active minus queued, capped by the remaining queue capacity.
A 503 provider_capacity_exceeded is separate and is retried with the same idempotency key.
import math
def wave_size(queue_capacity_remaining):
return max(1, math.floor(queue_capacity_remaining * 0.75))
for remaining in (5, 20, 100):
print(remaining, wave_size(remaining))Which error to expect
queue_full and rate_limited are both 429, so read the code, not the status. For a Free batch, six accepted jobs is the arithmetic, and for the provider side, see the 503 retry.
Sources
Related posts
More in Pricing
- Sume API requests per minute by plan: write and read limits
Writes per minute are 120 on Free, 300 on Pro, 600 on Startup and 1200 on Scale. Reads are 40 times that, so 4800, 12000, 24000 and 48000.
- Sume avatar quality defaults to plus when omitted: what it costs
Omit quality on an avatar video and Sume renders on Plus at $0.245 per second, 33 percent above Standard. How the default shows up in a 100-clip batch.
- Sume image markup at 10,000 images: list vs billed price per model
What the 1.25x house margin adds at 10,000 images per Sume image model, from $50 on Qwen Image and Grok Imagine to $527 on ChatGPT Image 2.
- Does Sume offer invoice or PO billing? What is public and what is not
The Sume pricing page lists invoice and PO billing under Enterprise with custom volume pricing. What is public, what is not, and how to budget before the call.
Written by Sume