Invisible models in Pika vs Sume Auto echoing sume/auto
Pika hides model names unless you ask. Sume's model sume/auto also hides the family and echoes sume/auto; pin a catalog id when you need a known model.

On the new Pika, models are named when you want them and invisible when you do not. Sume has the same two modes through the API: send model: "sume/auto" and Sume picks the family, or send a bare catalog id such as seedance-2 to pin one. With Auto, the response says sume/auto and does not tell you which family ran.
Pika's wording is from its September 17 post; Sume's from Video generation, read 2026-10-01.
What does Pika say about model choice?
The post says its apps automatically use the model best suited to your task, and that if you want more control you can still choose a specific model yourself. It names Seedance, GPT Image, MiniMax H3 and Pika's own models among those it uses.
What does Sume Auto return?
The Video generation page says model: "sume/auto" lets Sume pick the family, and the poll response reports "model": "sume/auto". It states that Sume does not disclose which family served the request, and that you should not build on any observable trait of the output to infer it. Resolution is a pure function of the normalized request and the catalog version, so an idempotent replay prices and routes the same way.
| You send | You get back | Use it when |
|---|---|---|
model: "sume/auto" | model: "sume/auto"; family not disclosed | You do not care which family renders |
| A bare catalog id | That id | You need a known model, duration range or reference type |
When should I pin a catalog id instead?
Pin when something downstream depends on the model: a 30-second length, reference video, or a budget you computed from one model's rate. The Video Router page's own guidance is "let Sume pick" for Auto and "pin a catalog model" for a chosen one. Auto create controls default to 720p and 8 seconds, so set resolution and duration explicitly if you need otherwise; see Auto defaults.
Can I tell which model made my clip?
Not from an Auto response. If you need to know, pin a catalog id from GET /v1/videos/models. For a comparison with another router's Auto, read Runway's media router vs Sume Auto.
Sources
Related posts
More in Models
- Pika Soundtrack: video in, sound out. What Sume has instead
Pika Soundtrack scores a finished video. Sume's docs offer generate_audio at video creation and a Timeline soundtrack bed, but no video-to-sound endpoint.
- 30-second video with sound: Pika Video Studio vs Sume ids
Pika Video Studio makes up to 30 seconds with sound. On Sume, seedance-2.5 and wan-3.0 reach 30 seconds; every other catalog model stops at 15.
- Pika Video Studio's 50 references vs Sume's per-model limits
Pika Video Studio attaches up to 50 references. Sume's caps are per model: 10 images and 3 short clips on Omni Flash, 1 to 8 images on Genjutsu.
- Qwen-Audio-3.0-TTS languages: 16, and Sume's language field
Qwen-Audio-3.0-TTS lists 16 languages. On Sume TTS you set the language field to a BCP-47 or ISO-639 code for every non-English transcript.
Written by Sume