AI spokesperson: build the avatar from a prompt, profile or photo?
Sume makes an avatar three ways: prompt, profile traits or a photo. Which one suits a brand spokesperson, what each costs, and the rights each one raises.
For a brand spokesperson who is not a real person, start with a prompt; use profile traits when your app has to produce avatars from data; use a photo only when the face belongs to someone who has agreed. All three cost the same flat $0.95 per avatar on Sume and give you the same thing: a reusable handle that every later video names.
The difference is control and rights, not price. The details below are from the Create new avatar docs and the OpenAPI reference, read 2026-10-03.
What does each input take?
Every request carries an avatar_handle, and an Idempotency-Key header makes a retry safe when you reuse it for the same payload. A leading @ is accepted and stored without it.
| Input | Request type | What you send | Best for |
|---|---|---|---|
| Prompt | prompt | Free text describing the avatar | A one-off brand character |
| Profile | props | ethnicity, sex (male or female), age 20-80 | Generating many avatars from your own data |
| Image | photo | A public HTTPS image_url | A real person who has signed a release |
Which input fits a brand spokesperson?
- A character you invent: use a prompt. You can write the look you want, then judge the result and create another handle if it is off.
- A cast you generate from a roster: use profile traits. The fixed fields make results easy to script, but the choice list is narrow: eight ethnicity values, two sex values.
- A named founder or employee: use a photo, with a written release first. See the release checklist.
- No wish to create anything: search the catalog of Sume's stock avatars, which use the reserved
sume_handle prefix, and render with one of those.
How do I keep the spokesperson consistent?
Consistency comes from reusing one handle. Every talking video names the handle, so the same face appears across a series, and each video holds one avatar. If you want different looks of the same person, Sume treats each look as its own avatar with its own handle; there is no look-pack feature, so plan the handles in advance, for example host_kitchen and host_office.
Before you commit to a spokesperson for a campaign, render a short check with the cheapest tier: a 4-second video at standard is under a dollar at $0.184 per second. A preview of first-frame stills is cheaper still to review, and it shows the face in the scene you plan to use.
What does a first test look like?
Create two avatars from a prompt, with different wording, and one from profile traits. Each is $0.95, so the whole test is $2.85. Render the same 5-second line with each at standard quality and compare. You are checking whether the face reads as a person your customers would believe in, whether it suits your product, and whether it holds up at phone size.
If none work, change the prompt rather than the tier. A higher quality tier changes the render, not who the person is. The avatar's identity is set when you create it, so a weak face is fixed at creation, not at render.
What can go wrong?
- Naming: a handle names one avatar, so give each variant its own clear handle.
- Prefix: handles starting with
sume_are reserved for Sume's own avatars and are rejected for yours. - Photo URL: the image_url must be public HTTPS, not a signed or private link.
- No delete route: the API lists create, list and read only, so create fewer, better-named avatars.
What about rights and disclosure?
An invented face has no signer, which removes the release step, but it does not remove the label. Viewers still need to know the spokesperson is AI where the platform or the law says so, and you still cannot present a generated person as a real customer. If the invented face happens to resemble a real person, you can still have a likeness problem, so look at the faces you generate before they run. Avatar 1.0 speaks English only in code, so a spokesperson for other markets needs a different route.
Sources
Related posts
- How to create a reusable AI avatar with the Sume Avatar 1.0 API
- Create an AI avatar from profile traits: the props input
- Create an AI avatar from a reference image: URL rules and cost
- Stock AI avatars API: find a ready-made avatar for a talking video
- AI avatar cost: $0.95 once to create, then per second of video
More in Sume Avatar 1.0
- How to check a face swap result: audio, length and stills
Check a Sume face swap before publishing: confirm the audio stream, compare length with the source, and sample stills with video inspect. Free probe.
- Does Sume face swap keep the original voice, outfit and scene?
Sume's Face Swap (Beta) keeps the source clip's audio, camera and timing but replaces the whole person, not just the face. What stays, what changes, cost.
- How to make an AI UGC ad look less staged with Avatar 1.0
Less-staged AI UGC comes from the first frame: phone-style framing, a casual scene prompt, an approved preview, then the final render. The levers Sume exposes.
- Shortest AI avatar video: 4 seconds, from $0.74 on Sume Standard
Sume avatar videos run 4 to 60 seconds. Per-second rates for Standard, Plus and Max, the 4 second floor, and what 15, 30 and 60 second clips cost.
Written by Sume