Omni replaces Veo in the Gemini app: does your API code change?
Google's page says Omni replaces Veo in the Gemini app. What that means for API code, and the one Omni row Sume lists, with its limits and per-second price.

No: nothing in Google's announcement changes an API call by itself, because the sentence that matters is scoped to a product. Google's Gemini Omni page says Omni "will replace Veo in the Gemini app", which describes what the app generates with, not which model ids an integration may send. If your code calls a video model through an API, the thing to check is the catalog of the API you call.
On Sume the answer is short. The published video catalog lists no Veo id, and it lists one Omni row, gemini-omni-flash-1.1. A pipeline that sends Sume that id keeps working the same way before and after Google's app change.
What Google's page says, and what it leaves out
The page is a consumer product page, so it describes the app. It also does not say what happens to existing Veo projects, and it gives no resolution. Treat anything beyond the lines below as unconfirmed.
| Topic | What the page states |
|---|---|
| Access | Google AI subscription required; features vary by tier and geography; 18+ |
| Length | Create 10 second videos |
| Audio | Native audio generation |
| Inputs | Turn photos into a video (up to 5) |
| Marked new | Scene extensions, video to video editing, multi-turn editing, avatar |
| Veo | Omni "will replace Veo in the Gemini app" |
| Resolution | Not stated on the page |
The Omni row you can call on Sume
Sume's video generation docs describe gemini-omni-flash-1.1 as 3 to 10 seconds at 360p, 720p, 1080p or 4K, in 16:9 or 9:16, with native synced audio that is always on. It takes text, a start frame and optional end frame, references (up to 10 images and up to 3 video clips of 3 seconds each), and a video_url for editing. It takes no audio references. Sume bills the provider list price times 1.25 per output second.
| Resolution | Provider list per second | Sume per second | 8-second clip |
|---|---|---|---|
| 360p | $0.03 | $0.0375 | $0.30 |
| 720p | $0.10 | $0.125 | $1.00 |
| 1080p | $0.15 | $0.1875 | $1.50 |
| 4K | $0.30 | $0.375 | $3.00 |
A short audit for your own code
- Search the repo for the model id string you send, not for the word Veo; the id is what a vendor retires.
- Read the live catalog with
GET /v1/videos/modelsbefore you hard-code a duration or resolution, since limits differ per model. - If a feature you need is only described on the app page, such as scene extensions, check the API docs for a field that does it before you plan around it. Sume's Omni docs list no extend input.
- Keep the id in config, so a swap is a one-line change and a re-run of your sample prompts.
Why the app and the API can diverge
An app picks a backend for you and can swap it overnight. An API gives you a named id, a documented envelope and a price per second. When a vendor says a model replaces another in its app, the useful question is whether a field you rely on, such as duration, resolution or a reference type, still exists in the API you call. For Omni on Sume those limits are the ones in the table above, and they are the same whatever the Gemini app does.
If your brief says Veo
A client brief may name Veo because that is the name they saw in the app. Rather than arguing about names, translate the brief into the properties it needs: clip length, aspect ratio, resolution, whether sound is included, and whether you must edit an existing clip. Then pick a catalog row that meets those properties. On Sume, an Omni row covers 3 to 10 second clips with sound in 16:9 or 9:16, plus an edit mode. If the brief needs something outside that, such as 15 seconds, the row table in Sume's docs shows which other ids reach it.
Sources
Related posts
More in Models
- Gemini Omni sound effects prompt: name each sound in 8 seconds
How to prompt footsteps, a door and breaking glass in a Gemini Omni clip: Google says describe audio explicitly. Sume request, timing and cost per take.
- Open-weights video news does not change your Sume bill: LTX, H3, Wan
LTX-2.5 and MiniMax H3 are reported open weights, but Sume still bills hosted per-second rows. Which ids to call and what each costs.
- Pick a video model for a Reel by length: 3, 10, 15 or 30 seconds
The length you need narrows the Sume models fast: 2 s is Wan only, 10 s allows Omni, 30 s allows Seedance 2.5 or Wan. Table of the shortlist.
- Portrait 4K AI video on Sume: Omni Flash 9:16 at $0.375 a second
LTX-2.5 is reported to do portrait up to 4K. On Sume the 4K portrait option is Gemini Omni Flash 1.1 at $0.375 a second: 3 s $1.13, 10 s $3.75.
Written by Sume