Agent 365 cost management for Copilot Studio agents vs a per-run cap
Agent 365 sets spend policy per user group and adds Copilot Studio agents in October. A per-run cap on the API call bounds the one task. Know which you need.

They answer different questions. Microsoft's Agent 365 cost management is a tenant-level control: policies about which AI models and capability levels different user groups can use, plus a consumption dashboard. A per-run cap, like Sume's generation_spend_cap_usd, is a call-level control for one task. If you run agents through Copilot Studio and also call outside APIs from them, you want both.
What Microsoft describes
The September post says Agent 365 cost management offers FinOps for AI, to manage usage-based spend, set spending guardrails, track costs and connect usage to value. It covers Copilot Cowork, WorkIQ, Code and Copilot Managed Runtime today, with agents built in Microsoft Copilot Studio scheduled for October. The dashboard shows usage trends, active users, credit consumption and estimated value. Microsoft says the capability is included with Microsoft cloud subscriptions.
Tenant policy versus call cap
Microsoft's post stresses that policies can restrict which models and capability levels groups may access, which is a good fit for keeping an expensive tier away from casual users. It does not claim to cap spend inside an external API called from an agent. Treat the dashboard as your view of Microsoft-side credits, and treat the API receipt as the source of truth for what the media provider charged.
| Question | Agent 365 cost management | Sume per-run cap |
|---|---|---|
| Scope of control | Groups of users, models and capability levels | One Agent Completion |
| When it applies | Policy set ahead of use; dashboard after | Before the run; no default, so the request fails without it |
| What it counts | Credits for Microsoft agent workloads | Generation spend on Sume |
| What it misses | Spend inside third-party APIs an agent calls | The agent's own language-model turn, billed to the separate Agent wallet |
The gap between the two
When a Copilot Studio agent calls a paid media API, Microsoft's dashboard shows the agent's credits and the API bills separately. Nothing joins the two unless you join them. Sume's receipt makes the join easy: every run has usage.billable_amount_usd_micros and generation_spend_cap_usd_micros, and the run id is stable as request_id in the signed webhook. Store that id next to the Microsoft-side record and you can reconcile later.
Because the Copilot Studio coverage is scheduled rather than shipped as of Microsoft's September post, confirm it is live in your tenant before you rely on it in a budget review.
What to set on the Sume side
On schedules, Sume's default cap is one dollar per run, and a per-run request value can only lower it. That keeps a cron mistake bounded even if nobody reads a dashboard.
generation_spend_cap_usdon everyPOST /v1/agent/completions, sized to one task.- An
Idempotency-Keyso retries return the original receipt withidempotency_hit: true. - A webhook with
outcomebranching, so adegradedrun (billed, no structured output) opens a review rather than a silent retry.
Sources
More in Pricing
- AI asset unit-cost sheet for client quotes: image, voice, video
A ten-row cost sheet from the Sume catalog: image, voice, transcript, cutout, music, render and video-second prices, all-in with the 5.5% fee.
- What do 50 AI image cards cost at low, medium and high quality?
Fifty Ideogram 4.5 images cost $1.875 at low, $3.75 at medium and $13.75 at high by Sume's catalog line. Read usage.cost on each result to confirm.
- AI video price per second, Oct 2026: Seedance, Wan, Kling, H3 on Sume
Billed price per second on Sume for Seedance 2.5, Wan 3.0, Kling 3 and MiniMax H3 at each resolution, with 5, 10 and 15 second clip totals.
- Audio edition of a weekly newsletter: 52 issues cost $9.88 on Sume
Narrating a weekly newsletter of 4,000 characters for 52 issues costs $9.88 on Sume TTS 1.0, $0.19 an issue. Set against MAI-Voice list prices.
Written by Sume