OpenAI Batch API limits vs Sume bulk runs: 50,000 vs 100
OpenAI Batch allows 50,000 requests per file with a 24h window. A Sume Format bulk run takes 1-100 items at concurrency 1-16, so chunk your file accordingly.

OpenAI's Batch guide allows up to 50,000 requests in one batch file. A Sume Format bulk run takes 1 to 100 items per queue, with concurrency from 1 to 16. A 50,000-line job becomes 500 queues of 100, and there is no 24-hour completion contract to plan around.
OpenAI figures are from its Batch guide; Sume's from Format bulk runs, both read 2026-09-30.
What are the limits side by side?
| Limit | OpenAI Batch | Sume bulk run |
|---|---|---|
| Items per submission | Up to 50,000 requests | 1 to 100 items |
| File size | Up to 200 MB | Not a file; a JSON body |
| In flight | Not stated as a setting | concurrency 1 to 16 |
| Completion window | 24h only | No window stated |
| Submission rate | 2,000 batches per hour | Not stated in the page |
How do I map a big file onto Sume?
Split the input into groups of at most 100 items, and send each as its own queue with a fresh Idempotency-Key; the docs warn that replaying a spent key returns 202 with the old queue. Each item has the same body as a single Format run. See Format bulk runs: 100 renders.
How do I read results?
There is no queue webhook; communication.webhook_url is per item, and queue progress comes from polling status_url. Queue completed means every item is terminal, not that all succeeded, so branch on counts.failed.
Can I cancel a queue?
Not as a queue: the docs say there is no public list-queues or cancel-queue endpoint. Cancel a child with POST /v1/format-runs/{run_id}/cancel.
Sources
Related posts
More in Developers
- OpenAI-compatible /v1/audio/speech: what Sume's TTS route is
Sume has no `/v1/audio/speech` clone. Its TTS is `POST /v1/tts-1.0/generate` or the Router, a job-based request authenticated with a Sume key.
- OpenAI whisper-1 shutdown Feb 2027: a Sume STT alternative
OpenAI removes whisper-1 on Feb 26, 2027. Sume's STT request names no provider model, takes a public audio URL, and caps reserved duration at 10 minutes.
- OpenAPI 3.2 generator: Sume's spec is still 3.0.3
Sume's published OpenAPI document declares 3.0.3, not 3.2. Check that your generator reads 3.0.x before pointing it at api.sume.com/reference/json.
- OpenCode MCP timeout: 5000 ms tools fetch vs Sume jobs_wait
OpenCode's remote MCP timeout is in milliseconds, default 5000, and covers fetching tools. Sume's jobs_wait holds up to 55 seconds per call, a separate limit.
Written by Sume