Avatar video quality settings: standard, plus or max?
Sume's talking avatar video takes quality standard, plus (default) or max. What each means, the other output fields, and how the preview relates to final tier.
Sume's avatar video takes quality: "standard" | "plus" | "max", and plus is the default when you omit it. standard is the fastest execution path, plus is the balanced path, and max is the highest tier with slower turnaround.
The three descriptions are all the Avatar video docs say, read 2026-09-29. They publish no visual comparison and no per-tier price in that page, so test on your own script and read the price from the API.
What do the tiers mean?
| `quality` | Docs description |
|---|---|
standard | Fastest Sume execution path |
plus | Default when omitted; balanced quality path |
max | Highest quality tier; slower turnaround |
What else shapes the output?
aspect_ratio:1:1,3:4,9:16(default),4:3or16:9.resolutionis currently720p.- The duration window is 4 to 60 seconds, from a
scriptor orderedvideo_inputs.
Can I preview before choosing a tier?
Yes. Avatar video previews (POST /v1/avatar-video-previews) produce still frames first. Preview stills are tier-independent, and you can override the final quality when you call /generate-video, so approving a preview does not lock in the tier. See first-frame previews.
How do I choose?
Start with the default plus for a draft. Move to max only for a script you will publish, and use standard when speed matters more, such as a bulk batch of tests. Read the reserve for each tier from the API's price fields before a big run; avatar video pricing per minute explains where to look.
How do I set it in a request?
Add quality next to avatar_handle and one of script or video_inputs.
curl -X POST https://api.sume.com/v1/avatar-1.0/talking-video \
-H "Authorization: Bearer $SUME_API_KEY" \
-H "Content-Type: application/json" \
-H "Idempotency-Key: avatar-max-001" \
-d '{
"avatar_handle": "your_avatar_handle",
"aspect_ratio": "9:16",
"quality": "max",
"script": "Here is the one setting that matters."
}'Sources
Related posts
More in Developers
- C# text to speech: call a TTS API with HttpClient
Text to speech in C#: POST the text and a voice with HttpClient, poll the job until it finishes, then stream the MP3 from its audio_url to a file.
- Text to speech streaming API: what Sume returns instead
Sume's text to speech API doesn't stream audio chunks. It returns a finished file per job; split long scripts into sentence jobs to start playback sooner.
- Duck background music under a voiceover with the Timeline API
Set soundtrack.duck_db (0 to 20) on POST /v1/timeline-1.0/render so the music dips under your voiceover spine. It needs a real spine; silence mode is refused.
- Render a silent video from clips with the Timeline API
Set audio.mode to silence and a duration_seconds on POST /v1/timeline-1.0/render to join clips with no audio file. Which fields are illegal there, and pricing.
Written by Sume