Case study video narrated by an avatar: three scenes, 55 seconds
Narrate a written case study as a 55-second, three-scene Sume avatar video: problem, change, result. Cost by tier and a body you can send; no names needed.
A written case study becomes a 55-second Sume avatar video in three scenes (the problem, what changed, the result), and it costs $10.12 on Standard, $13.48 on Plus and $30.25 on Max. The narrator is a generated presenter, not the customer, so the video should read as a summary of the written study, with every claim traceable to it.
The video_inputs shape and the 4-60 second window are from Sume's avatar video guide, read 2026-10-04.
Three scenes from a one-page study
Take one sentence from each section of the written case study and turn it into spoken lines. Do not invent a quote or attribute a statement to a person who did not say it.
| Scene | Seconds | Source in the written study |
|---|---|---|
| problem | 15 | The situation before |
| change | 20 | What the team did |
| result | 20 | The measured outcome |
Cost by tier
| Item | Standard | Plus | Max |
|---|---|---|---|
| One video | $10.12 | $13.48 | $30.25 |
| Four case studies | $40.48 | $53.90 | $121.00 |
Request body
Spoken scenes use voice.type: "text"; the total planned duration must stay between 4 and 60 seconds, and scene backgrounds must resolve to one shared scene.
curl -X POST https://api.sume.com/v1/avatar-1.0/talking-video \
-H "Authorization: Bearer $SUME_API_KEY" \
-H "Content-Type: application/json" \
-H "Idempotency-Key: case-study-001" \
-d '{"avatar_handle": "narrator", "aspect_ratio": "16:9", "quality": "plus", "video_inputs": [{"id": "problem", "voice": {"type": "text", "script": "The support team handled every refund request by hand, and replies took days.", "duration": 15}, "background": {"type": "prompt", "prompt": "Bright office, neutral wall"}}, {"id": "change", "voice": {"type": "text", "script": "They moved the first reply into a short video that explains the refund steps.", "duration": 20}, "background": {"type": "prompt", "prompt": "Bright office, neutral wall"}}, {"id": "result", "voice": {"type": "text", "script": "Repeat questions dropped, and the team now uses the saved time for harder cases.", "duration": 20}, "background": {"type": "prompt", "prompt": "Bright office, neutral wall"}}]}'Claims, approval and honesty
Give the customer the final script and keep their approval. Say a number only if it is in the written study, and say it as the study states it. Say in the clip, the caption or the page around it that the presenter is AI-generated; the API does not add that label for you.
Sources
Related posts
More in Use cases
- Cyber Week five-day creative sprint: a daily Seedance clip budget
Shopify defines BFCM as Thanksgiving through Cyber Monday. Twenty-two 8-second 720p vertical clips across those five days cost $101.6928 on Sume.
- Demand Gen carousels: 2 to 10 matching cards from one reference
Demand Gen carousels take 2 to 10 cards. Image assets run 4:5 or 9:16 at 5 MB. How to batch matching cards from one reference image with the Sume image API.
- Demand Gen logo 150 KB, images 5 MB: check sizes before upload
Demand Gen allows images up to 5 MB but logos only 150 KB. Check bytes on every Sume image before upload, since output compression is not served in v1.
- Demand Gen video ads: four ratios, a 5 s floor and the 4:5 gap
Demand Gen video takes 1:1, 16:9, 4:5 and 9:16, from 5 s. How to render the four-ratio set with Sume, including the 4:5 case the video API does not list.
Written by Sume