Does an image cost more at 16:9 than 1:1 on Sume? Flat rows vs GPT 2.5
On most Sume image rows the ratio does not change the price. ChatGPT Image 2.5 is the exception: cost follows pixels and quality. Billed figures for both.

On most Sume image models, changing aspect_ratio does not change the price: Seedream, Flux, Qwen, Imagen, Grok and Recraft each have one per-image rate in the repo's rate table, and Nano Banana and Ideogram 4.5 price by tier or quality, not by ratio. ChatGPT Image 2.5 is the exception, because it is billed from output tokens that depend on pixels and quality.
Flat per-image rows
| Model | Billed per image | What moves the price |
|---|---|---|
| Seedream 4.5 | $0.05 | Nothing in the ratio |
| Flux 2 Pro | $0.0375 | Nothing in the ratio |
| Qwen Image | $0.025 | Nothing in the ratio |
| Nano Banana 2.1 | $0.10 at 1K | Tier (0.5K $0.075, 2K $0.15, 4K $0.20) |
| Ideogram 4.5 | $0.0375 low, $0.075 medium, $0.275 high | Quality; fal lists the same price for all sizes |
ChatGPT Image 2.5 follows pixels
For ChatGPT Image 2.5, Sume estimates the output with OpenAI's size and quality calculator and charges fal's $30 per million output tokens as list, then times 1.25. The same quality costs less on a smaller canvas. Figures below are for high quality, output tokens only.
| Canvas | List | Billed |
|---|---|---|
| 1024x1024 (1:1) | $0.0527 | $0.065875 |
| 1536x864 (16:9) | $0.0324 | $0.0405 |
| 720x1280 (9:16) | $0.0285 | $0.035625 |
| 2560x1440 (16:9, 2K) | $0.0553 | $0.069125 |
What to take from it
- A 16:9 thumbnail on GPT 2.5 at 1536x864 is cheaper than a 1024 square at the same quality, because it has fewer pixels' worth of tokens in this estimate.
- On the flat rows, do not squeeze a banner into a square to save money. The price is the same.
- On GPT 2.5,
autoquality reserves themaxprice before the job runs, so set quality yourself when you budget. - Input images on edits add input tokens on top of the output estimate.
Budget formula
For flat rows, budget is images times the billed rate. For GPT 2.5, estimate per canvas and quality. Forty squares at 1024 and high quality cost 40 x $0.065875 = $2.635, and at low quality 40 x $0.007375 = $0.295. Edits that send reference images add input tokens.
Do not reuse a high-quality 1024 figure for other sizes or qualities. If a spend cap matters, set quality explicitly and check usage.cost on the first call.
Check your own call
Read usage.cost from the response, or the pricing line from GET /v1/images/models/{id}/endpoints. The catalog list for GPT rows is the high-quality 1024 square; other sizes come from the estimator at submit time. The Image API page describes the token rates, and OpenAI's image generation guide lists the custom-size rules that GPT sizes follow.
Sources
Related posts
More in Pricing
- Draft 20 Wan 3.0 clips at 480p, then finish the best six at 1080p
Twenty 5-second Wan 3.0 drafts at 480p cost $6.40; finishing six at 1080p adds $7.50. All twenty at 1080p cost $25.00.
- Price per kept AI clip when 1 in 3 is rejected: Wan 720p
A 15-second Wan 3.0 clip at 720p is $1.875 on Sume. If you keep one clip in three, the real price per kept clip is $5.625. The keep rate is a budget input.
- 8-second 720p Seedance price ladder: mini, fast, 2.0 and 2.5 on Sume
One 8-second 720p clip costs about $1.51 on seedance-2-mini and $4.62 on seedance-2.5. The token formula behind each Sume price, with the 1.25 margin shown.
- ElevenLabs v4 voice cost vs a 30-second avatar clip: under 1%
ElevenLabs v4 lists $0.08 per 1,000 characters. A 30-second script is about 450 characters, or 3.6 cents, against $7.35 for a plus-tier Sume avatar clip.
Written by Sume