Google Ads built-in image and video generation vs a Sume pipeline
Demand Gen can generate up to 20 images per prompt and auto-build video from a logo, two images and two texts. Where a separate Sume pipeline differs.

Google Ads can generate images from a prompt inside Demand Gen, with up to 20 selectable results per prompt and a choice of aspect ratio, and can auto-generate video in horizontal, square and vertical from a logo, two images, two generic texts and a business name. A separate Sume pipeline gives you per-request model and ratio control and Sume-hosted result files, at the cost of one more step before upload.
What Google's tools take as input
The Google Ads Help page on generative AI in Google Ads describes two features. Facts below were read on 2026-10-03.
| Feature | What the page says |
|---|---|
| Image generation | Generate images from prompts; select up to 20 per prompt; choose the aspect ratio |
| Video generation | Auto-generates video in horizontal, square and vertical |
| Video inputs required | 1 logo, 2 images, 2 generic texts, 1 business name |
Where a Sume pipeline differs
Both routes can produce usable assets. The differences are in what you control and where the files live. The Sume facts below come from the Image API and video docs.
| Question | Google Ads built-in | Sume pipeline |
|---|---|---|
| Choose the model | Not a request field on the page read | Pass a catalog model id, or sume/auto to let Sume pick |
| Aspect ratio | Choose per prompt for images | aspect_ratio per request, limited to the model's descriptors |
| Images per call | Up to 20 selectable per prompt | n up to 10 per call; per-model ceilings are lower |
| Video length | Not stated on the page read | Per model: most to 15 s, some to 30 s; Timeline assembles longer |
| Files | Not stated on the page read | Sume-hosted, signed URLs |
| Retries | Not stated | Idempotency-Key makes a replay return the original job on video |
When each makes sense
Use the built-in tools when speed inside the account matters more than control, such as a quick test of a Demand Gen group from existing brand images. Use a separate pipeline when you need the same character or product across ratios, a fixed model for comparison tests, or the same assets reused on other channels.
A caveat for sume/auto: the docs state that responses echo sume/auto and the family that ran is never disclosed, so pin a model id when a test needs one fixed model.
A minimal Sume side
A video request for a vertical placement looks like this. It follows the documented POST /v1/videos shape, with a retry-safe header.
curl -X POST https://api.sume.com/v1/videos \
-H "Authorization: Bearer $SUME_API_KEY" \
-H "Content-Type: application/json" \
-H "Idempotency-Key: dg-hero-001" \
-d '{
"model": "sume/auto",
"prompt": "A vertical product clip on a desk, natural light, slow push-in",
"aspect_ratio": "9:16",
"duration": 5
}'What this does not decide
Neither route settles ad policy, disclosure, or whether a synthetic asset suits your brand. Those checks apply to the finished creative either way, so run the same review on both.
Sources
Related posts
More in Comparisons
- Image reference limits: 10 on ElevenLabs, 16 on Sume, 5 Ideogram 4.5
ElevenLabs lists 10 references for GPT Image 2.5. Sume's docs list 16 for the same models and 5 total for Ideogram 4.5. A table, a count guard and an edit rule.
- Kling 4.0 vs Kling 3.0: the spec differences in one dated table
Length, resolution, HDR, references, keyframes, audio and prompt size, Kling 4.0 against 3.0 as stated on Kling's pages, plus how each maps to a Sume request.
- Luma API callbacks and credit balance vs Sume webhooks and /v1/balance
Luma's docs list callbacks and a credits balance. Sume has webhook mode, GET /v1/balance, and a generation_limits snapshot. How to use each before a batch.
- Luma Ray 2 API parameters: keyframes, loop and callback_url
Luma's video docs list ray-2 and ray-flash-2 with keyframes, loop, concepts and callback_url. Here is each field mapped to the Sume /v1/videos request.
Written by Sume