GPT Image 2.5 edit chains: four high passes cost one max image
At 1024x1024, four high-quality GPT Image 2.5 passes add up to the output price of one max image. What that means for a one-change-per-pass edit chain on Sume.

A four-pass edit chain at high quality costs the same output price as one max image at 1024×1024 before input tokens and Sume pricing: 4 × $0.05268 = $0.21072. fal's guide and Sume's docs both give $0.21072 for max, so the choice between one big prompt at max and four small passes is a real trade, not a rounding detail.
Numbers come from fal's How To Use GPT Image 2.5 (read 2026-10-02) and Sume's Image API page, which uses the same Fal token rates. Edit advice comes from OpenAI's Image prompting guide. Your billed amount is the usage.cost in each response.
Why would anyone chain passes?
Both guides recommend it. fal says one change per pass, because combining several edits in one instruction raises the chance of a retry. OpenAI says to pass the previous output as the next edit input, request one change and repeat the details to preserve. The same guides warn that repeated edits can still shift details you meant to keep, so each pass needs its constraint sentence.
What does a chain cost at 1024×1024?
These are output-image prices only at 1024×1024. Input tokens (references) and Sume's pricing are added on top, and the figures are estimates from the vendors' calculator.
| Quality | Per image | Four passes |
|---|---|---|
| low | $0.00588 | $0.02352 |
| high | $0.05268 | $0.21072 |
| xhigh | $0.09366 | $0.37464 |
| max | $0.21072 | $0.84288 |
How do I keep the cost down?
Run the early passes cheaply and the last one at the quality you ship. A pass at low is about one ninth of a high one. If a pass is not what you wanted, only a completed image is billed; a failed generation is not. Sume's docs also note that auto quality reserves max, so set a quality explicitly when you budget.
- Pick the edit order so the riskiest change comes first.
- Use
lowormediumwhile you test wording. - Finish at
highor above only once. - Sum
usage.costin your own code.
How do I sum a chain on Sume?
This script applies one instruction per pass, feeds each result back as the next reference, and prints the running usage.cost. If Sume rejects a result URL as a reference, save a copy and re-host it on a public HTTPS address. It needs pip install requests.
import os, requests
STEPS = ["Brighten the sky.", "Remove the bin on the left.",
"Warm the colour grade slightly."]
KEEP = "Keep everything else exactly as it is."
url = "https://example.com/photo.jpg"
total = 0.0
for i, step in enumerate(STEPS):
quality = "high" if i == len(STEPS) - 1 else "low"
r = requests.post(
"https://api.sume.com/v1/images",
headers={"Authorization": f"Bearer {os.environ['SUME_API_KEY']}"},
json={"model": "openai/gpt-image-2.5",
"prompt": f"Image 1 is the photo to edit. {step} {KEEP}",
"quality": quality,
"input_references": [
{"type": "image_url", "image_url": {"url": url}}]},
timeout=60,
)
r.raise_for_status()
body = r.json()
url = body["data"][0]["url"]
total += body["usage"]["cost"]
print(i + 1, quality, url, f"running cost {total:.4f}")When is one max pass the better choice?
When the instruction is a single clear change and the quality of fine detail matters most, one max image can be cheaper in your time than a chain, because there are fewer places for drift. When you are exploring wording, the chain at low is far cheaper. The prices above do not decide it alone; the number of retries does.
Whichever you choose, set the quality explicitly. Sume's docs say an auto quality reserves max, and an omitted value defaults to high.
Sources
- Image API
- [How To Use GPT Image 2.5: Prompts & Workflows [2026] (fal, read 2026-10-02)](https://fal.ai/learn/tools/how-to-use-gpt-image-2-5)
- Image prompting (OpenAI, read 2026-10-02)
Related posts
More in Pricing
- grok-imagine-image-2.0 price: $0.04 to $0.08 per image, $0.01 input
xAI lists grok-imagine-image-2.0 from $0.04 (1K, low) to $0.08 (2K, medium) per output image, plus $0.01 per input image on edits. Table inside.
- How many 5-second AI video clips does $100 buy?
A $100 wallet buys 1,600 five-second Grok Imagine clips, 160 Wan 3.0 clips at 720p or 142 Kling 3 clips. Computed from the Sume video catalog, 2026-10-01.
- How much money can Sume hold in flight per plan?
Accepted job capacity is 6 on Free up to 120 on Scale. At $2.50 per 10-second Wan 1080p job, a full Scale queue holds $300. Computed from the docs.
- Kling motion control duration_seconds: it sets the credit hold
On Sume's Kling 3.0 Motion Control, duration_seconds (1 to 30) sizes the credit reservation; the output length follows the driving video, not this field.
Written by Sume