9:16 Vertical AI Video: Which Models Support It, and Sume Auto Limits
Which video models output 9:16 for Reels, Shorts and TikTok: Veo 3.1, Gemini Omni, Kling 3, Seedance on Sume, plus what sume/auto accepts.

Vertical is the default delivery format for short video, but not every model renders it. Some only offer 16:9 and 9:16, some add 1:1 and 21:9, and cropping a landscape clip to vertical throws away most of the frame. Here is what each documents on 2026-10-03.
Support by model
Google's Veo and Omni docs both list 16:9 and 9:16. On Sume, kling-3 takes 16:9, 9:16 and 1:1; the Seedance 2 family lists a wide set from 21:9 down to 9:16, with 4:3, 1:1 and 3:4 in between; and gemini-omni-flash-1.1 takes 16:9 or 9:16.
| Model | Ratios documented |
|---|---|
| Veo 3.1 (Google) | 16:9, 9:16 |
| Gemini Omni Flash 1.1 | 16:9, 9:16 |
| kling-3 on Sume | 16:9, 9:16, 1:1 |
| seedance-2 on Sume | 21:9, 16:9, 4:3, 1:1, 3:4, 9:16 |
| sume/auto | 16:9 or 9:16 |
What sume/auto accepts
Sume's Auto route, model: "sume/auto" on POST /v1/videos, defaults to 720p and 8 seconds and accepts 3 to 10 seconds at 16:9 or 9:16. That covers the common vertical ad. It does not cover 1:1 or 21:9, and it will not run a clip longer than 10 seconds; the docs also say Sume never discloses which family served a request, so pin a model when you need a specific one.
Pass the ratio explicitly. A vertical prompt with no aspect_ratio is not a vertical request.
Delivery notes
A 9:16 clip at 720p is 720 by 1280 pixels, below the 1080 by 1920 most platforms ask for. If the final needs full vertical HD, either pick a model that renders 1080p natively or generate at 720p and upscale. Sume's Timeline renders 1080 by 1920 by default, so a 720p source gets resampled when assembled.
For model-by-model selection beyond ratio, see the five questions to ask before pinning one.
Sources
Related posts
More in Models
- 9:16 vertical AI video: which Sume models take an aspect ratio
Seedance, Wan 3.0, Kling 3, MiniMax H3 and Gemini Omni Flash take 9:16; Grok Imagine, Genjutsu and H3 Max Recast take no aspect ratio. A per-model table.
- AI video ratios on Sume: Omni Flash is 16:9 or 9:16, Kling adds 1:1
Gemini Omni Flash 1.1 on Sume takes only 16:9 and 9:16; Kling 3 adds 1:1. Which other Sume video models list more ratios, and which list none.
- Video-to-video clip length limits on Sume, by tool
Recast takes 5 to 30 s, Genjutsu 4 to 30 s, Omni edit 3 to 10 s, Avatar face swap about 4 to 15 s. A table and a Python picker choose by clip length.
- Vietnamese, Thai, Indonesian, Malay TTS API: vi, th, id, ms on Sume
Cartesia Sonic 3.6 lists vi, th, id and ms. How to send each as the language on Sume TTS, why the field is required, and a per-character cost for each script.
Written by Sume