Speechmatics TTS $0.011 per 1,000 characters vs Sume TTS limits
Speechmatics lists TTS at $0.011 per 1,000 characters with the first million free. A full 20,000-character Sume request would cost $0.22 at that rate.

Speechmatics' pricing page lists text-to-speech at $0.011 per 1,000 characters, with the first 1 million characters free. At that rate a 20,000-character script costs $0.22, and the first 50 such scripts would be free. Sume TTS accepts up to 20,000 characters per request but publishes no TTS rate in its docs, so measure a real job on your balance.
The Speechmatics line
Speechmatics' page is the source for the rate and free allowance.
| Item | Figure | Worked example |
|---|---|---|
| Rate | $0.011 per 1,000 characters | 20,000 characters = $0.22 |
| Free allowance | First 1 million characters | 50 full 20,000-character scripts |
| Sume request cap | 20,000 characters | One request, one job |
Sume TTS request
Choose a voice from a Sume avatar or voice.id, set language for non-English text, and pick a container. Audio over 1,200 seconds fails and is not charged.
import os, requests
r = requests.post(
"https://api.sume.com/v1/tts-1.0/generate",
headers={"Authorization": f"Bearer {os.environ['SUME_API_KEY']}",
"Idempotency-Key": "tts-demo-001"},
json={"transcript": "Welcome back. Today we cover three updates.",
"avatar_id": os.environ["SUME_AVATAR_ID"],
"generation_config": {"speed": 0.95, "emotion": "warm"},
"output_format": {"container": "mp3", "sample_rate": 44100,
"bit_rate": 128000}},
)
print(r.status_code, r.json())Making the comparison honest
Quality differs by voice, language and script, which a price table cannot show. Generate the same three paragraphs on each service, then compare the charge and the sound.
Worked example
The free million characters go further than you might expect.
- 20,000 characters at $0.011 per 1,000 is $0.22.
- 1 million characters is 50 such requests, free on Speechmatics' page.
- 100 requests after the free allowance would cost $22.00 in characters alone.
Checklist before you commit
Price is one input; voice choice and language support matter just as much, so listen before deciding.
- Check language coverage.
- Check output formats.
- Re-read the pricing page before launch.
Sources
Related posts
More in Comparisons
- Stable Audio DAW plugin: BPM sync and takes vs Sume Music
Stability's Stable Audio plugin runs in Logic Pro and Ableton Live with BPM sync. Sume has no plugin or tempo field; Music Router returns an audio file by API.
- Stable Audio web app mixing vs Sume timeline soundtrack gain and duck
Stable Audio's web app has level, pan, mute and solo per track. Sume's timeline soundtrack has gain, ducking, loop and fade-out, but no pan, mute or solo.
- StepAudio 3 Gen: voice, SFX and music in one clip vs Sume jobs
StepFun's stepaudio-3-gen-preview makes voice, effects, ambience and music in one audio output. Sume uses separate music, speech and Timeline mix steps.
- StepAudio 3 Music from a dry vocal or reference audio vs Sume
StepAudio 3 Music accepts lyrics, vocals or reference audio, even scoring a dry vocal. Sume Music takes a text prompt and one optional image, no audio.
Written by Sume