5 product photo variations from one packshot: GPT 2.5 vs Seedream
Five scene variations of one packshot cost $0.10 on GPT Image 2.5 medium with one reference, $0.25 on Seedream 4.5. Request shape and cost table on Sume.

Five variations of one product packshot cost about $0.104 on GPT Image 2.5 at medium quality with one reference at 1024x1024, and $0.25 on Seedream 4.5 at its flat $0.05 per image. Both take input_references, so the flow is the same: pass your packshot URL, describe the new scene, set n. The Image API docs describe the reference format and limits.
The surprise is the order. The newer model is cheaper here because medium GPT Image 2.5 is billed by tokens, and one reference adds an estimated input-token charge of well under a cent.
What does the request look like?
The packshot must be a public HTTPS URL; Sume rejects localhost, private-network and non-HTTPS URLs before it submits. Four variations in one call, one more in a second:
{
"model": "openai/gpt-image-2.5",
"prompt": "Same bottle, unchanged label, on a sunlit bathroom shelf with eucalyptus",
"input_references": [
{"type": "image_url", "image_url": {"url": "https://example.com/packshot.png"}}
],
"image_size": {"width": 1024, "height": 1024},
"quality": "medium",
"n": 4
}How do the two models price five variations?
GPT Image 2.5 reserves an estimate that includes the input tokens of your reference, which Sume's catalog describes as an estimate rather than a measured count. Seedream is a flat rate per output image.
| Model and tier | Per image | Five images |
|---|---|---|
| GPT Image 2.5, medium, 1 reference | $0.0209 | $0.104 |
| Flux 2 Pro | $0.0375 | $0.1875 |
| Seedream 4.5 | $0.050 | $0.250 |
| GPT Image 2.5, high, 1 reference | $0.0835 | $0.4175 |
Which should I pick?
Run the same packshot and prompt on both for one image each, then compare label fidelity, because a changed label is the failure that costs you. Choose on that result, not on price. Prices here are list x 1.25 in dollars per image before whole-cent rounding; the catalog's billable formula reads "list x 1.25, ceil usd cents", so confirm the first charge in usage.cost and budget from the invoice, not from the sum.
Sources
Related posts
More in Comparisons
- Groq Orpheus TTS: $22 per million characters and a 200-character cap
Groq lists Orpheus V1 English at $22.00 per million characters with input kept under 200 characters. Request-count and cost math against Sume TTS.
- HeyGen Pro 4K export at $49 vs Sume video upscale per clip
HeyGen's page puts 4K export on Pro at $49 a month and 1080p on Creator at $29. Sume Video Upscale charges $0.009 per input second. Here is the arithmetic.
- How much more does 4K AI video cost than 720p? Multiplier table
Veo 3.1 Fast 4K costs 3 times its 720p rate on Google's page, Omni on Sume is also 3 times, Veo Standard 1.5 times and Kling Video v3 Pro 1 times.
- Ideogram 4.5 is 0.8c to 22c per image: which tiers does Sume sell?
Ideogram lists four quality modes from 0.8 cents to 22 cents at native 2K. Sume's Image API sells low, medium and high at $0.0375, $0.075 and $0.275 an image.
Written by Sume