OpenAI Agents API vs a custom agent API for async runs

OpenAI's Agents API keeps durable sessions; Sume Agent Completions return a 202 receipt to poll or receive by webhook. Where each fits, and how they differ.

4 min readSume
All posts

They solve different problems, so the choice is about what you need managed. OpenAI's Agents API manages sessions, orchestration, context compaction, and recovery for an agent you configure. Sume's Agent Completions run the Sume Agent, with its sandbox, tools and media generation, on a task you send each time and return a 202 receipt you poll.

What is each one, in the vendors' words?

OpenAI's overview describes an OpenAI-managed API where your application provides tools and chooses its execution environment. Its changelog lists the public beta on 2026-09-10.

Sume's docs describe an Agent Completion as running the Sume Agent on an ad-hoc prompt, the same runtime as the Agents chat UI, reachable from your backend with an API key.

How do the two differ on the points that decide it?

This table only restates each page; it does not rank them.

Agent run APIs side by side, read 2026-09-29.
PointOpenAI Agents APISume Agent Completions
StatusPublic betaDocumented as shipped, with a Not available yet list
SessionsDurableEach completion runs in a fresh thread
Learning the resultStream or webhooksPoll the receipt or an agent.run.terminal webhook; no streaming yet
ToolsConfigured at session setupSume Agent tools, MCP bridge, media generation

Is a Sume completion a chat completion?

No. The docs say it is not a synchronous chat completion, because a real agent turn opens a sandbox and may generate media. The create call returns 202 with a receipt. The request borrows OpenAI's messages[] shape, but the response is a run receipt, not choices[].

What must a Sume run set up front?

You must set a ceiling per run: generation_spend_cap_usd has no default, and omitting it fails the request with 400 invalid_request. Size it with the metered rates on the API pricing page. OpenAI's overview says its API bills at model rates plus tool and container fees.

Which one fits which job?

If you want OpenAI to manage the session and you bring your own tools, OpenAI's page describes that. If you want to hand a task to the Sume Agent and get back generated media with durable media.sume.com URLs, Sume's Agent Completions do that, and Formats fit when only the inputs change between runs.

Neither excludes the other. Read the Not available yet list on Sume's page and OpenAI's beta notes before you commit a workflow to either.

What does polling look like?

Poll GET /v1/agent-runs/{id} with your key. Statuses are queued, processing, completed, failed and canceled. A completed run fills output, by default the sume/action-run-output/v1 shape with the closing text in output.text, and POST /v1/agent-runs/{id}/cancel stops a run in flight.

Sources

Related posts

More in Developers

All Developers posts

Written by Sume