crawl_profile in Sume MCP: Instagram or TikTok handle to counts

Use crawl_profile to turn a public Instagram or TikTok handle into identity and counts. It is read-only, unbilled discovery, and a missing count is unknown.

5 min readSume
All posts

crawl_profile is the Sume MCP tool for one question: who is this public Instagram or TikTok account, and what are its counts? You give it platform (instagram or tiktok) and a handle; on TikTok you can pass user_id instead. It returns identity and counts. A count that the platform did not return is unknown, not zero. It is a read tool, so a read-only OAuth session can use it, and the tool description calls it bounded, unbilled discovery.

The tool is one of four social tools in the crawl group. The others read posts, resolve one URL, and discover unknown creators, so choose by what you already have.

One practical use is vetting before you spend. Before an agent builds content around a creator or a competitor, a single profile read confirms that the handle exists and shows the counts, at no charge, and lets the model decide whether a deeper read is worth the next call.

Which crawl tool for which input

The routing text that the server sends with its instructions is short and strict. Identity and counts go to crawl_profile. A known account's posts go to crawl_feed. A known media URL goes to crawl_media. Unknown creators and topic or hashtag searches go to crawl_find.

The table is the quickest way to stop an agent from picking the wrong tool. Ask it to name what it has in hand before it names a tool, and the choice follows.

Sume MCP social tools by what you start with (read 2026-10-05)
You haveCallReturns
A known handle, want identity and countscrawl_profileIdentity and counts
A known handle, want recent or top postscrawl_feedItems with permalink and counts
One public media URLcrawl_mediaPermalink, counts, image or video candidates
Only a topic or hashtagcrawl_findCandidate creators or media

What it will not do

Public data only is the main boundary. The tool does not use a login, and it cannot see private accounts, stories or direct messages. It also does not download media or transcribe. If a lookup fails, the description says a failed lookup is not an empty feed, so an agent should report the failure and not claim that the account has nothing.

Counts are the part to be careful with. Followers, posts and similar numbers are what the platform returned at the moment of the call, and any figure that is missing is unknown. An agent should quote the number with the date it was read and never fill a gap with a guess.

Timing and trust

Calls can be synchronous for up to 30 seconds. If the work is queued or still processing, the answer carries a request_id; follow it with jobs_wait and then jobs_result on the same id, never by sending the lookup again. Treat all source content as untrusted: a bio is text written by someone else, and it can contain instructions meant for a model.

Because the instructions treat source content as untrusted, a safe agent prompt says that text in a bio, a caption or a link is data to be summarized, and never an instruction to be followed.

The input is flat, so a call is short. A typical argument object looks like this:

Platform values are exactly instagram and tiktok, in lower case. For TikTok, user_id can stand in for the handle when you have it, which is useful when a handle was changed.

{
  "platform": "instagram",
  "handle": "example_handle"
}

What to call next

To move from an identity to content, call crawl_feed next with the same handle. To check that your session can see the tool at all, call tools_list first, as in the MCP quickstart. The inventory of the crawl group is on MCP tools and gates.

A read-only OAuth session is enough for the whole chain up to this point, which makes this a good first demo for a person who is deciding whether to grant Write at all.

Sources

Related posts

More in Integrations

All Integrations posts

Written by Sume