12 Sume video ids: which need a source video, which take text
Nine of Sume's twelve video ids take a text prompt. grok-imagine-video-1.5 needs an image; higgsfield-genjutsu and h3-max-recast need a video. Ranges per id.

Sume's video router lists twelve ids. Nine start from a text prompt. grok-imagine-video-1.5 needs a first-frame image, and higgsfield-genjutsu and h3-max-recast need a source video plus photos, so none of those three can start from a prompt alone. gemini-omni-flash-1.1 additionally has an edit mode that takes a video_url.
The twelve
Durations come from the docs and the model catalog; confirm them live with GET /v1/videos/models.
| Id | Duration | Starts from |
|---|---|---|
| seedance-2.5 | 4 to 30 s | Text, images, references |
| seedance-2 | 4 to 15 s | Text, images, references |
| seedance-2-fast | 4 to 15 s | Text, images, references |
| seedance-2-mini | 4 to 15 s | Text, images, references |
| wan-3.0 | 2 to 30 s | Text, images, references |
| kling-3 | 4 to 15 s | Text, first and last frame |
| grok-imagine-video-1.5 | 4 to 15 s | A first-frame image (no text-only) |
| minimax-h3 | 5 to 15 s | Text, images, references |
| minimax-h3-max | 5 to 15 s | Text, first/last frame, references |
| gemini-omni-flash-1.1 | 3 to 10 s | Text, image, references, or a video to edit |
| higgsfield-genjutsu | 4 to 30 s | One source video plus 1 to 8 images |
| h3-max-recast | 5 to 30 s | Source video plus 1 to 4 person photos |
Why it matters
If you wrap the API in a picker, filter on the request you can actually build. A form with only a prompt box should hide grok-imagine-video-1.5, higgsfield-genjutsu and h3-max-recast; the catalog constraints say each requires an image or a source video, so a prompt alone will not produce a clip.
- Prompt-only forms: show the nine text-capable ids.
- Tools that upload footage: add
h3-max-recastandhiggsfield-genjutsu; tools that upload a still can addgrok-imagine-video-1.5. - Always read
supported_input_referencesfrom the catalog instead of hard-coding this table.
Reading the catalog in code
The live source for this table is the model list. A short script can print every id with its durations and input types, which makes a good CI check when the catalog changes:
import os, requests
r = requests.get("https://api.sume.com/v1/videos/models",
headers={"Authorization": "Bearer " + os.environ["SUME_API_KEY"]})
for m in r.json()["data"]:
d = m["supported_durations"]
print(m["id"], d[0], d[-1], m["supported_input_references"])Sources
Related posts
More in Models
- Veo 3.1 clip older than 2 days: it cannot be extended
Veo 3.1 extends only Veo-made clips kept 2 days. If yours expired, here is what Google's page allows and what Sume's Omni edit and new generation do.
- When Seedance 2 Mini or Fast is enough, and when to pay for 2.5
A decision guide for Seedance tiers on Sume: what Mini, Fast and 2.5 cost for the same clip, what only 2.5 can do, and which jobs belong on which tier.
- Which video model for a 12 s 9:16 clip with end frame and sound?
Seedance 2.5, Wan 3.0, Kling 3 and MiniMax H3 all take a 12-second 9:16 clip with an end frame on Sume. Prices run from $0.90 to $6.94; here is the pick.
- Which Sume video models stop at 15 seconds and which reach 30?
Seedance 2.5, Wan 3.0, Genjutsu and H3 Max Recast go past 15 seconds on Sume; MiniMax H3 stops at 15 and Gemini Omni Flash 1.1 at 10. The full limit table.
Written by Sume