Claude MCP connector or Sume Agent Completions for a long video job?
A Messages API request with the MCP connector holds open while Claude works. Sume Agent Completions return 202 and a run to poll. Which one fits a long job.

Use Sume Agent Completions when the job is long and nobody is watching, and use Claude's MCP connector when a Claude conversation should drive Sume tools turn by turn. Agent Completions return 202 and a receipt that you poll, with a required spend cap. The MCP connector is a beta feature of the Claude Messages API that lets Claude call tools on a remote MCP server, so your code runs a conversation and carries the Claude side of the cost.
The two shapes
The two shapes differ in who runs the agent loop, as the table shows (read 2026-10-08).
| Question | Claude MCP connector | Sume Agent Completions |
|---|---|---|
| Status | Beta, header mcp-client-2025-11-20 | Shipped; POST /v1/agent/completions |
| Who runs the agent | Claude, in your Messages request | The Sume agent in Sume's runtime |
| Server reachability | Public HTTPS, Streamable HTTP or SSE | Not applicable |
| MCP features used | Tool calls only | Sume's own tools and sandbox |
| Result | Messages response | 202 receipt, then poll /v1/agent-runs |
| Spend control | Your own; Sume max_spend_usd optional | generation_spend_cap_usd required |
| Streaming | Per Messages API | Not available |
Why long jobs favor the receipt
A render can outlast any single request. With the connector you manage the wait: call jobs_wait in slices up to 55 seconds, keep the conversation alive, and avoid a duplicate paid create when a request times out. With Agent Completions, the create call returns at once and you poll status_url until next_action is not poll_status. Statuses are queued, processing, completed, failed and canceled. You can send a communication.webhook_url to be told at a terminal status.
The connector fits when the model must reason between tool calls, for example choosing among catalog options a user is discussing. Anthropic's page also notes the connector is not ZDR-eligible and is not available on Bedrock or Google Cloud.
A minimal Agent Completion
The key must carry agent_completions:write to create and agent_completions:read to poll; an older key without them gets 403 insufficient_scope. Service-account keys cannot create completions.
curl -sS -X POST https://api.sume.com/v1/agent/completions \
-H "Authorization: Bearer $SUME_API_KEY" \
-H "Idempotency-Key: render-2026-10-08-001" \
-H "Content-Type: application/json" \
-d '{"instruction": "Make a 10 second product teaser.", "generation_spend_cap_usd": 5}'
# then poll
curl -sS https://api.sume.com/v1/agent-runs/$RUN_ID \
-H "Authorization: Bearer $SUME_API_KEY"Decision rules
The shortest rule is who owns the loop. If your code and your model should own the plan, use the connector or your own MCP client. If Sume's agent should own the plan and you want a receipt, use Agent Completions.
The second rule is attendance. A person in a chat can approve spend, so a connector-driven conversation can ask first. An unattended backend cannot, which is why Sume makes the cap mandatory on the completion route and why you should send max_spend_usd on every paid MCP call.
- Interactive and tool-by-tool: connector or MCP client with
dry_runfirst. - Unattended, one task per call: Agent Completions with a spend cap and webhook.
- Saved workflow with changing inputs: a Format run, not a completion.
Sources
Related posts
More in Comparisons
- Custom avatars per plan: HeyGen, Colossyan, Tavus, Teams, Sume
Custom avatars each plan lists: HeyGen 1 to 10+, Colossyan 15 or 20, Tavus 0 to 7, Teams 3. Sume prices each new avatar at $0.95; no plan cap is listed.
- flux-2-pro vs flux-2-flex on Sume: price, limits and a 100-image bill
flux-2-pro is $0.0375 and flux-2-flex is $0.0625 per image on Sume. Both take 10 references and n up to 4. Compare the 100 and 1,000 image bill.
- Four Sume image rows between 3.75 and 5 cents: Flux, Seedream, Recraft
Flux 2 Pro $0.0375, Seedream 5.0 Lite $0.04375, Seedream 4.5 and Recraft V4 $0.05. Compare ratios, references and formats before you choose on price alone.
- Haiku 5.5 vs Sonnet 5.5 token prices for an agent calling Sume
Haiku 5.5 costs $0.10 per million input tokens and Sonnet 5.5 costs $2. What the gap does and does not change for a Sume agent's spend cap.
Written by Sume