ElevenLabs Dubbing v2 cloning_strength 0-10: what Sume has instead

ElevenLabs Dubbing v2 has cloning_strength from 0 to 10, default 7. Sume has no such dial: you pick a voice per language and set speed. What that costs you.

5 min readSume
All posts

ElevenLabs Dubbing v2 has a cloning_strength setting from 0 to 10 with a default of 7. Higher values favour resemblance to the original speaker; lower values give the model more freedom for natural delivery in the target language (ElevenLabs dubbing docs, read 2026-10-02). Sume has no equivalent dial: you choose a library voice and the language, and text to speech reads your translated script.

What the ElevenLabs page says

The docs label v2 as Alpha with automatic processing and no in-app transcript editor, and v1 as in maintenance mode with Dubbing Studio. Its dubbing API blog (read 2026-10-02) says Dubbing v2 does not adjust lips: audio lands where the original dialogue did through sync-aware translation.

ElevenLabs dubbing, read 2026-10-02
ItemWhat the docs say
cloning_strength0 to 10, default 7
Higher valuePrioritises similarity to the original speaker
Lower valueMore natural delivery in the target language
v2 languages90+ with regional dialects for some, such as en-AU
File limits1 GB and 180 minutes in the app; 3 GB per file via the API for v2
SpeakersUp to 32 unique speakers per file
Self-serve concurrencyUp to 3 concurrent dubbing jobs
Background audioKept, so music and effects are not re-mixed

What you control on Sume

On Sume you set the voice, the language and the text. Text to speech also takes speed, and sentence segments let you fit each line to its slot, as covered in Translate a video to another language with AI voice, on time. The voice does not shift toward the original speaker by a dial; it is simply the voice you picked.

Background audio is yours to keep: detach the original track, and mix the new speech over it with Timeline audio parts.

The trade-off

Which is better depends on whether the original speaker's identity matters more than control.

  • ElevenLabs trades some naturalness for resemblance at high cloning_strength; Sume has no resemblance setting at all.
  • ElevenLabs v2 handles speaker separation; on Sume you separate speakers yourself.
  • Sume gives step-level control and per-step pricing; ElevenLabs gives one project per dub.

Sources

Related posts

More in Comparisons

All Comparisons posts

Written by Sume