HeyGen Video Agent edit_plan vs Sume preview regenerate
HeyGen's edit_plan revises named scenes in one turn, up to 50 items. Sume has no in-place scene edit: regenerate preview stills, then render.

HeyGen's Video Agent now accepts an edit_plan on POST /v3/video-agents/{session_id}: one natural-language change per scene, applied together in a single turn, with untouched scenes keeping their script and visuals. The Sume docs describe no equivalent in-place scene edit. Its closest flow is to regenerate the preview stills, approve them, then start the final render.
HeyGen facts are from its September 2026 changelog; Sume facts are from Avatar video previews and Generate avatar video, read 2026-09-30.
How does edit_plan work?
You read scene ids and edit_version from GET /v3/videos/{video_id}/scenes, then send items with scene_id, text, scene_snapshot_video_id and edit_version. A scene outside the snapshot returns 400 invalid_parameter; a snapshot that changed since you read it returns 409 stale_edit_version. Up to 50 items per request.
What does Sume offer instead?
Avatar video previews generate the first-frame stage without the full talking-video render, so you can approve composition before spending on it. POST /v1/avatar-video-previews/:id/regenerate reuses the stored request and refreshes only the stills. POST /v1/avatar-video-previews/:id/generate-video then starts the render, reusing the preview first frame when available.
| Step | HeyGen Video Agent | Sume avatar video |
|---|---|---|
| Change one scene | edit_plan item per scene | Regenerate stills for the preview |
| Conflict handling | 409 stale_edit_version | Structural changes need a new preview |
| Multi-scene input | Scenes read from the video | Ordered video_inputs list |
Which changes need a new preview on Sume?
The docs say structural fields (script, video_inputs, avatar_handle, scene, aspect_ratio) still require a new preview. Only the final render tier can change at the generate step. Note that regenerate refreshes stills for the whole stored request, not one chosen scene.
How should I plan edits on Sume?
Put each scene in its own entry of video_inputs, review scene_previews[], and start a new preview when the script changes. For the wider editing picture, see edit an AI avatar video.
Sources
Related posts
More in Developers
- HeyGen video scenes API vs Sume job result previews
HeyGen's GET /v3/videos/{video_id}/scenes returns each scene's visuals and script. Sume job results return media.sume.com artifacts and scene previews.
- Hookdeck 15-minute timeout and Sume's 10-second webhook attempt
Hookdeck's longer destination timeout does not change Sume's fixed 10-second attempt. Ack fast at the relay URL and give your handler its own time budget.
- How long does Sume retry a webhook before giving up?
Sume job webhooks give up after about 4.5 minutes of gaps, run webhooks after about 3 hours. Cumulative timings per attempt, then redeliver or poll.
- x402 vs Sume's 402: a prepaid-balance error, not a payment prompt
Sume returns HTTP 402 with the code insufficient_credits when the balance is too low. The fix is adding funds, not sending a payment header.
Written by Sume