Runway Product Ad recipe vs Sume reference-to-video
Runway's Product Ad Recipe turns product photos into an ad video. On Sume, send photos as reference images and put the concept in one prompt.

Runway's Product Ad Recipe takes one to ten product photos plus optional style images, productInfo and a userConcept, and returns a finished ad video. On Sume there is no storyboard step: you send the photos as reference images to a video model that accepts them, and write the concept, product facts and camera direction into one prompt.
Recipe fields are from Runway's Product Ad page. Sume details are from the Video Router docs and Video generation, read 2026-10-01.
How do the Product Ad fields map?
Runway's page says the Recipe analyzes the product, builds a storyboard and renders a short video. Sume's closest shape is reference-to-video.
| Runway field | Runway limit | Sume side |
|---|---|---|
productImages | 1 to 10 images | reference_image_urls (up to 10) on gemini-omni-flash-1.1 |
styleImages | Up to 4 | Same list, described in the prompt |
productInfo / userConcept | 2,500 / 3,500 characters | One prompt |
duration | 4 to 15 seconds | 3 to 10 seconds on that model |
audio | Default false | Native audio always on |
How do I point the prompt at a photo?
With reference_to_video, address media as <IMAGE_REF_0>, <IMAGE_REF_1> and so on, zero-based in list order. Write "the bottle in <IMAGE_REF_0> on a marble counter, slow dolly in" so the model knows which image is the product. On /v1/videos the equivalent field is input_references, which gives style or content guidance rather than exact frames.
Which models accept references?
Only models whose supported_input_references lists a type accept that type, and limits are per model. Read capabilities from GET /v1/video-router/models rather than assuming one envelope. Photos must be reachable over public HTTPS. Following Runway's own tip is also useful on Sume: clean, high-resolution product photos from several angles give the model more to work from.
What should I do next?
Start with one short clip and two or three angles of the product. For a luxury look see AI luxury product ad; if you already have an ad and only the product changes, use an edit instead, as in the product swap comparison.
Sources
Related posts
More in Use cases
- AI presenter for a SaaS demo video, with or without a product image
For a SaaS demo, omit `product_image` for a productless avatar video, or add one. Each Sume avatar job runs 4 to 60 seconds, so longer demos are split.
- AI video background swap for a fashion lookbook with Seedance 2.5
Seedance 2.5 describes green screen editing that swaps a background and keeps the subject. On Sume, send the garment clip as a video reference, 4 to 30 seconds.
- Seedance 2.5 green screen editing: what Sume can do
ByteDance says Seedance 2.5 improves green screen editing. Sume's docs document an edit mode only for Gemini Omni Flash 1.1; Seedance takes video references.
- AI workout video generator: Seedance 2.5, 4 to 30 seconds
For a workout demo, one Seedance 2.5 request runs 4 to 30 seconds on Sume, with video and audio references; most other catalog models stop at 15 seconds.
Written by Sume