Gemini 3.8 Flash TTS batch: $4.50 per 1M audio tokens vs $9.00
Batch halves Gemini 3.8 Flash TTS output to $4.50 per 1M tokens through 2026, then $9.00 in 2027. Sume TTS has one flat per-character rate.

Yes: the Gemini API pricing page lists Gemini 3.8 Flash TTS batch output at $4.50 per 1M audio tokens through Dec 31 2026, against $9.00 on the standard tier. Both double on Jan 1 2027. Sume TTS has no batch tier; it bills one per-character rate for every job.
The Google price grid
All figures are per 1M tokens from Google's pricing page. Input is text; output is audio.
| Model and tier | Input (text) | Output (audio) to Dec 31 2026 | Output from Jan 1 2027 |
|---|---|---|---|
| Flash, standard | $0.50 | $9.00 | $18.00 |
| Flash, batch | $0.25 | $4.50 | $9.00 |
| Flash-Lite, standard | $0.50 | $6.00 | $12.00 |
| Flash-Lite, batch | $0.25 | $3.00 | $6.00 |
What the batch discount does not tell you
The price is per token, not per character or per minute. The model page lists Batch API support, and the tokens page we read gives an audio rate for input only, so we cannot convert $4.50 into a cost per minute of speech. Run a sample script through the API and divide what Google reports by the seconds you got.
Batch also changes delivery: you trade immediate output for a lower rate, so it suits overnight ad-variant runs, not an editor waiting on a take.
How Sume prices the same job
Sume TTS 1.0 charges $0.0475 per 1,000 characters (see the API reference and GET /v1/catalog). Spaces and punctuation count, and each job rounds up to the next cent with a one-cent minimum. A 600-character voiceover is about 2.9 cents at list rate before rounding up.
Sume does not offer a discounted batch lane. If you submit many jobs you pay the same rate per job, but you can cap spend per request with max_spend_usd and preview with dry_run.
Which to pick
For a very large library re-voiced once, measure Gemini batch on your real scripts first. For small, repeated ad reads where you want a known cost before you run, a per-character rate is easier to forecast.
Sources
Related posts
More in Pricing
- Gemini 3.8 Flash TTS context caching vs Sume transcripts
Gemini 3.8 Flash TTS prices cached input at $0.125 per million tokens until 2026-12-31. Sume sends the full transcript on each job. When caching matters.
- Gemini 3.8 Flash TTS: batch and flex $4.50, priority $16.20 per 1M
Gemini 3.8 Flash TTS is $9 per 1M audio tokens Standard, $4.50 batch or flex, $16.20 priority until Dec 31, then double. Cost of 1,000 minutes.
- Gemini Omni 4K is $0.375 a second, double 1080p: worth $3.75 a clip?
Omni Flash 1.1 on Sume costs $0.375 a second at 4K, twice 1080p and three times 720p. A 10-second 4K clip is $3.75 against $1.25 at 720p. When it pays.
- GPT Image 2.5 at 4:5 lists 12-14% below 1:1 at the same quality
On Sume, GPT Image 2.5 at 4:5 (1024x1280) lists 12 to 14 percent below a 1024x1024 square at every quality tier. The five-tier table, with billed prices.
Written by Sume