Nano Banana 2.1 or GPT Image 2.5 for product photos on Sume
A feature matrix for product shots on Sume: references, ratios, mask_url, transparent background, tiers and price for Nano Banana 2.1 and GPT Image 2.5.

For product photos on Sume, pick GPT Image 2.5 when you need a transparent cutout, a mask or more than 10 references, and Nano Banana 2.1 when you want a fixed price per tier and a wider ratio list. Both take reference images, and both are called through the same POST /v1/images.
How do the two rows differ?
The first four rows come from the Sume catalog code and docs on main. The last row is the vendors' own pricing pages.
| Feature | Nano Banana 2.1 | GPT Image 2.5 (Flare, Sunburst) |
|---|---|---|
| References on Sume | up to 10 | up to 16 |
| Ratios on Sume | 14 plus auto | 8 plus auto |
| mask_url | no | yes |
| background transparent | no | yes |
| Size control | 512, 1K, 2K, 4K tiers | quality tiers; custom sizes in multiples of 16 |
| Vendor image output rate | $30.00 per 1M tokens (Google) | $30.00 per 1M tokens (OpenAI) |
Which one for a white-background catalog shot?
Either. Ask for a plain white backdrop and give the product photo as an input_references item. Nano Banana 2.1 is simpler to budget: $0.10, $0.15 or $0.20 billed by tier. GPT Image 2.5 bills by quality and size, so run the endpoint line for your settings first.
Which one for a transparent cutout?
GPT Image 2.5. Set background: "transparent" and a format that carries alpha. Nano Banana 2.1 rejects background with 400 unsupported_parameter.
{
"model": "openai/gpt-image-2.5",
"prompt": "Isolated wireless earbuds case, studio lighting",
"background": "transparent",
"output_format": "png",
"aspect_ratio": "1:1"
}What does GPT Image 2.5 cost?
The image models docs give a 1024 by 1024 example: xhigh output is $0.09366 and max is $0.21072 at list, before input tokens and Sume pricing. Lower qualities cost less. The Sume billed price adds the 1.25 factor, so read the endpoint line rather than adding it yourself.
A fair test costs little. Use the same product photo, the same plain prompt and a 1:1 frame, run each model once, and compare label text, edge quality and how well the product shape is kept. Keep the reference photo sharp and evenly lit, because both models copy its flaws. Decide on that sample, not on the matrix alone.
Sources
Related posts
More in Comparisons
- Nano Banana 2.1 vs GPT Image 2.5 on Sume: 4.0x at 1K, 0.9x at 4K
Sume lists Nano Banana 2.1 at $0.10, $0.15 and $0.20 and GPT Image 2.5 at $0.02475, $0.055625 and $0.2225. The price ratio flips below 1 at 4K.
- Nine prices for one 5-second 720p AI video, from 25 cents to $2.00
A 5-second 720p clip on Luma Ray 3.2, LTX-2.5, MiniMax H3, Runway Gen-4.5, Veo 3.1 and Sume rows, with each vendor's own arithmetic. Cheapest to dearest.
- Deepgram Nova-3 batch $0.26/hour vs streaming $0.29: 500 hours vs Sume
Deepgram lists Nova-3 monolingual at $0.0043 a minute pre-recorded and $0.0048 streaming. Over 500 hours that is $129 against $144; Sume STT is $300.
- One provider behind Hy Image 3.5 Preview vs Sume's one image endpoint
OpenRouter forwards Hy Image 3.5 Preview to one provider, Tencent Cloud. Sume's image API serves every model from one endpoint. What each means for routing.
Written by Sume