12 reference photos in one edit: only GPT Image 2.5 takes them

Nano Banana 2.1 takes 10 references on Sume, Ideogram 4.5 takes 5, GPT Image 2.5 takes 16. For a 12-photo edit use GPT Image 2.5, or split the job in two.

5 min readSume
All posts

If one edit needs 12 reference photos, openai/gpt-image-2.5 is the Sume row that takes them: its input_references range is 0 to 16. Nano Banana 2.1 stops at 10, and Ideogram 4.5 at 5. Tencent's Hy Image 3.5 Preview lists up to 20 on OpenRouter, but Sume does not list it.

Limits side by side

Sume's shared reference ceiling is 10, with per-model exceptions of 16 for both ChatGPT Image 2.5 ids and 5 for Ideogram 4.5. Imagen 4 takes none.

Reference limits, Sume catalog and OpenRouter page, as of 2026-10-08
ModelMax referencesWhere
openai/gpt-image-2.516Sume catalog
google/nano-banana-2.110Sume catalog
ideogram/ideogram-v4.55Sume catalog
google/imagen-4-fast, imagen-4-ultra0Sume catalog
Hy Image 3.5 Preview20OpenRouter page; not on Sume

The price catch

GPT Image 2.5 is billed on tokens, and each reference adds input tokens. The docs give $8 per million input image tokens at the provider and say input counts are estimates. With no size or quality set, a text-to-image call quotes $0.2225 on Sume and a one-reference edit $0.28175. Set image_size and quality to keep a 12-photo edit predictable.

Splitting instead

If 10 references are enough once you pick the best ones, use Nano Banana 2.1 at $0.10 for 1K. Rank your photos, keep the 10 that show the product label, color and shape, and drop duplicates. Fewer, cleaner references usually beat a pile of near-identical ones, but test on your product before you trust that.

Choosing the references

Sort the 12 photos by what each shows: two for the overall shape, two for the label, two for color accuracy, two for texture and four for context. If you must cut to 10, drop one from context and one from texture first. The label and the shape matter most.

If you go with GPT Image 2.5, also pass mask_url when you want to limit the edit to part of the image, which that row lists. Nano Banana 2.1 does not list a mask field, and a request with one returns a 400.

Compare on one test product before you commit a catalog: run the 12-reference GPT edit and a 10-reference Nano Banana 2.1 edit on the same item and look at the label.

What the numbers rely on

Reference limits come from Sume's catalog code: a shared ceiling of 10, with 16 for both GPT Image 2.5 ids and 5 for Ideogram 4.5. The Hy Image figure of 20 is from OpenRouter's model page. Treat 20 as the OpenRouter listing, not as a Sume limit.

The code can change. Read supported_parameters.input_references from the live catalog.

If your photos are public on a product page, you can pass those URLs directly. If they sit behind a login, host copies on a public HTTPS path first. Sume rejects localhost, private-network and non-HTTPS URLs before submission, so a failed submit on a good-looking URL is usually a private host. Test one URL with a single edit call before you queue a whole set.

Sources

Related posts

More in Comparisons

All Comparisons posts

Written by Sume