Gemini TTS API: which engine Sume's TTS routes to

Gemini 3.8 TTS is not in Sume's TTS Router, which lists Cartesia Sonic ids only. What Sume's TTS takes for voice, engine and length, with a checklist.

4 min readSume
All posts

Sume does not route to Gemini TTS. The Sume TTS Router catalog is Sonic only: model is an enum of sonic-3.6, sonic-3.5, sonic-3, sonic-latest and sonic-preview, and sume/tts-1.0 has no engine picker at all. If your plan depends on a Gemini model, call Google's API; if it depends on a voice and a finished audio file, read the checklist below.

Gemini facts are from Google's launch post for Gemini 3.8 Flash TTS and Flash-Lite TTS. Sume facts are from the API reference OpenAPI description, read 2026-09-30.

What does Google say Gemini 3.8 TTS offers?

Google's post lists new voices created from natural-language prompts, 2,000+ production-ready voices, voice replication from a 30-second sample, two-speaker scene staging, and more than 100 languages and dialects. It says every clip is watermarked with SynthID, and that voice replication requires a verbal consent recording. Availability is the Gemini API, Google AI Studio, Gemini Notebook and Google Vids; the article does not give pricing.

What does Sume's TTS take instead?

A Sume TTS request takes a transcript and a voice selector: a top-level avatar_id or avatar_handle, or voice.id. The reference says the discoverable route is the avatar list, GET /v1/avatar-1.0/avatars, using an avatar whose voice.status is ready. There is no field for describing a voice in words.

Sume's TTS is an asynchronous job: you submit, then poll or receive a webhook. Synthesized audio longer than 1200 seconds fails with tts_duration_exceeded, and no credit is captured.

Checklist from Google's launch post and the Sume API reference, read 2026-09-30.
QuestionGemini 3.8 TTS (Google)Sume TTS
Voice from a text promptYes, per the postNo field for it
Pick a voice2,000+ ready voicesavatar_id, avatar_handle or voice.id
Pick an engineGemini 3.8 Flash or Flash-LiteTTS Router model: Sonic ids only
Length capNot stated in the post1200 s per job
DeliveryGemini API and Google appsAsync job, poll or webhook

How do I pick a Sonic model on Sume?

Send the Sonic id in model to POST /v1/tts-router/generate. The list of ids comes from GET /v1/tts-router/models; an unknown id fails with 400 model_not_found.

curl -X POST https://api.sume.com/v1/tts-router/generate \
  -H "Authorization: Bearer $SUME_API_KEY" \
  -H "Content-Type: application/json" \
  -H "Idempotency-Key: gemini-checklist-001" \
  -d '{
    "model": "sonic-3.6",
    "transcript": "Your order ships tomorrow.",
    "avatar_handle": "acme",
    "language": "en"
  }'

What should I do if I need Gemini voices?

Use Google's API for the Gemini audio, then bring the file into your pipeline. Sume's model ids and what each Sonic version changes are in Cartesia Sonic 3.6 API model ids, and the general request shape is in the text-to-speech API guide.

Sources

Related posts

More in Models

All Models posts

Written by Sume