Flux TTS expressivity -2 to 2 vs Sume's emotion guide
Deepgram's Flux TTS expressivity runs -2 to 2 (0 nominal). Sume TTS has no such dial: generation_config takes volume, speed and a free-text emotion guide.

Deepgram's Flux TTS has a new expressivity control from -2 to 2, where 0 is nominal. Sume TTS 1.0 has no numeric equivalent: generation_config takes volume, speed and an optional emotion guide that is a short string. To get a flatter or more animated read on Sume, describe it in emotion and compare takes.
Deepgram facts are from its changelog entry for the Aug 26, 2026 self-hosted release; Sume facts from the OpenAPI schema and the API reference, read 2026-10-01.
What does Flux TTS expressivity do?
The entry says the control makes a Flux TTS voice more calm or more animated. Negative values are flatter, positive values more animated, and 0 is the nominal setting. It appears in a self-hosted release note, so check where your deployment runs before relying on it.
What does Sume offer instead?
Three optional controls grouped under generation_config. Only two are numeric, and neither is an expressivity scale.
| Control | Deepgram Flux TTS | Sume TTS 1.0 |
|---|---|---|
| Expressivity dial | -2 to 2, 0 nominal | None |
| Emotion | Not in the entry | emotion: string guide, 1 to 64 characters |
| Speed | Not in the entry | speed: multiplier in [0.6, 1.5] |
| Volume | Not in the entry | volume: multiplier in [0.5, 2.0] |
How do I get a calmer or livelier read on Sume?
Pass a plain description such as calm or excited in emotion; the schema calls it an optional guide, so treat it as a suggestion and listen. Change one thing at a time, and use speed for pacing rather than mood. More on the three fields in text to speech with emotion, speed and volume.
const body = {
transcript: "The launch moved to Thursday.",
avatar_handle: "@your_avatar",
generation_config: { emotion: "calm", speed: 0.9, volume: 1.0 },
mode: "async",
};
console.log(JSON.stringify(body));Can I map -2 to 2 onto emotion strings?
Only by your own convention, since Sume documents no scale. A table such as -2 as flat, 0 as no emotion field and 2 as animated keeps your app's dial stable, but the same word may land differently on different voices.
Sources
Related posts
More in Models
- Gemini 3.1 flash image preview shutdown: Sume model ids
Google lists gemini-3.1-flash-image-preview and gemini-3-pro-image-preview for June 25, 2026 shutdown. Sume callers use catalog ids like google/nano-banana-2.
- GPT Image aspect ratios for a Pinterest pin (2:3) and Instagram (4:5)
Sume lists 17 aspect ratios, including 2:3 for a Pinterest pin and 4:5 for an Instagram portrait. A model only accepts the ratios its catalog row lists.
- HappyHorse 1.0 API: what Runway lists and how to check Sume's models
Runway lists HappyHorse 1.0 at 3-15 seconds, 720p and 1080p. Sume's video docs name no such id, so read /v1/videos/models before you plan around it.
- Hedra Character 3 aspect ratios vs Sume Avatar Video ratios
Hedra Character 3 lists seven aspect ratios including 9:21 and 21:9. Sume Avatar Video supports five: 1:1, 3:4, 9:16, 4:3 and 16:9, default 9:16.
Written by Sume