MCP tool search: how a long Sume tool list loads in Claude Code
Claude Code loads MCP tools on demand with tool search, which is on by default. What that means for Sume's long hosted tool list and how to prompt for it.

Claude Code's MCP page says tool search is on by default: rather than putting every MCP tool description into the model's context, it loads them when the model needs them. That suits Sume's hosted server, whose tool list is long, because a video task only needs a handful of tools.
The page was read 2026-09-29; the Sume tool names come from its MCP docs.
What does Claude Code's MCP page say?
A few details matter for a large server.
| Topic | Detail |
|---|---|
| Tool search | On by default |
| Tool names | mcp__<server>__<tool> |
| Output limit | MAX_MCP_OUTPUT_TOKENS, default 25,000 tokens, warning at 10,000 |
| Scopes | local, project (.mcp.json), user |
Which Sume tools does a video task need?
Most of the tool list is for other jobs: the crawl_* family, assets_*, timeline_* and more. A video task usually touches five.
tools_listandtools_schemato discover and inspect a contractgenerate_videoto submit, withpayload.modelomitted to route tosume/autojobs_waitto wait on the returned job idsjobs_resultto read the finished media URLsbalance_getorgeneration_admission_previewto check funds first
How do I help the model find them?
Name the tools in the prompt. With tool search the model discovers tools by need, and a prompt that says which Sume tool to use for which step removes guesswork. Have it call tools_schema with a tool name before its first paid call so it reads the exact contract, including idempotency_key and dry_run.
You can also pre-approve a short list in Claude Code's permissions instead of leaving the whole server open; the site's page on allowing MCP tools shows the rule syntax.
What if a result is too large?
Sume's job results are short JSON with media URLs, so they sit well under Claude Code's default 25,000 token output limit. A crawl or catalog listing can be larger; ask for a narrower call rather than raising the limit. To do many calls in one step, script_run runs a short JavaScript program on Sume's side with sume.call and returns a calls[] journal and jobs[].
Sources
Related posts
More in Agents
- Run the Sume video agent from your backend with Agent Completions
POST /v1/agent/completions runs the same agent as the Sume Agents chat, with tools and media generation, and returns an async run receipt you poll or webhook.
- Safe automation for AI agents that call paid APIs
Keep agents read-only by default, keep secrets out of logs, and on hosted MCP send an idempotency_key, preview with dry_run, and cap with max_spend_usd.
- Scheduled AI video agent runs: cron, API triggers, and receipts
A Sume schedule is a saved Agents automation that runs on a cron cadence and returns a run receipt. Author it in the dashboard; start and monitor runs by API.
- What is a video agent? How Sume defines and runs one
In Sume's docs, a video agent is a sandbox Agent that composes generation tools into a post-ready video. Brief it in chat, or call it over HTTP.
Written by Sume