OpenAI Batch API limits vs Sume bulk runs: 50,000 vs 100
OpenAI Batch allows 50,000 requests per file with a 24h window. A Sume Format bulk run takes 1-100 items at concurrency 1-16, so chunk your file accordingly.

OpenAI's Batch guide allows up to 50,000 requests in one batch file. A Sume Format bulk run takes 1 to 100 items per queue, with concurrency from 1 to 16. A 50,000-line job becomes 500 queues of 100, and there is no 24-hour completion contract to plan around.
OpenAI figures are from its Batch guide; Sume's from Format bulk runs, both read 2026-09-30.
What are the limits side by side?
| Limit | OpenAI Batch | Sume bulk run |
|---|---|---|
| Items per submission | Up to 50,000 requests | 1 to 100 items |
| File size | Up to 200 MB | Not a file; a JSON body |
| In flight | Not stated as a setting | concurrency 1 to 16 |
| Completion window | 24h only | No window stated |
| Submission rate | 2,000 batches per hour | Not stated in the page |
How do I map a big file onto Sume?
Split the input into groups of at most 100 items, and send each as its own queue with a fresh Idempotency-Key; the docs warn that replaying a spent key returns 202 with the old queue. Each item has the same body as a single Format run. See Format bulk runs: 100 renders.
How do I read results?
There is no queue webhook; communication.webhook_url is per item, and queue progress comes from polling status_url. Queue completed means every item is terminal, not that all succeeded, so branch on counts.failed.
Can I cancel a queue?
Not as a queue: the docs say there is no public list-queues or cancel-queue endpoint. Cancel a child with POST /v1/format-runs/{run_id}/cancel.
Sources
Related posts
More in Developers
- OpenAI whisper-1 shutdown Feb 2027: a Sume STT alternative
OpenAI removes whisper-1 on Feb 26, 2027. Sume's STT request names no provider model, takes a public audio URL, and caps reserved duration at 10 minutes.
- OpenCode MCP timeout: 5000 ms tools fetch vs Sume jobs_wait
OpenCode's remote MCP timeout is in milliseconds, default 5000, and covers fetching tools. Sume's jobs_wait holds up to 55 seconds per call, a separate limit.
- OpenRouter video provider.options on Sume: rejected, not dropped
OpenRouter lists provider passthrough configuration. Sume v1 runs one backend per model, so a non-empty provider.options returns 400 unsupported_parameter.
- OpenRouter video seed on Sume: no v1 model accepts it
OpenRouter lists seed for deterministic video generation. On Sume no v1 video model accepts seed: each reports seed false and the field is rejected.
Written by Sume