GPT Image 2.5 cached input price: does Sume pass it on?
OpenAI lists cached input at $1.25 text and $2 image per million tokens for Sunburst. Sume's docs give estimates and reserved cost, not a cache discount.

OpenAI's Sunburst page lists cached input at $1.25 per million text tokens and $2 per million image tokens, but Sume's docs describe no cache discount. Budget from the endpoint's pricing lines and the estimates in the docs, not from the cached rate.
OpenAI's figures come from its model pages; Sume's from the Image API docs. Read 2026-09-30.
What does OpenAI list?
The Sunburst page lists text input at $5 (cached $1.25), image input at $8 (cached $2), and image output at $30 per million tokens. The Flare page lists the same $5, $8 and $30 rates and says "Cached inputs receive 75% discounts".
What do Sume's docs say?
Sume states that Flare and Sunburst use the same Fal token rates: $30 per million output image tokens, $8 per million input image tokens, and $5 per million input text tokens. Output estimates use OpenAI's ChatGPT Image 2.5 size and quality calculator. At 1024x1024, xhigh output is $0.09366 and max output is $0.21072 before input tokens and Sume pricing. Input token counts are estimates, and Fal rounds the total up to $0.0001.
The docs mention no cached-input rate. They also say auto quality reserves max.
| Line | OpenAI (Sunburst) | Sume docs |
|---|---|---|
| Text input | $5 (cached $1.25) | $5 |
| Image input | $8 (cached $2) | $8 |
| Image output | $30 | $30 |
| Cache discount | Listed | Not described |
What should I budget on?
Endpoint pricing lines are the amount charged to your wallet, with Sume's margin already applied, so cost_usd x n is what you pay. Read them from GET /v1/images/models/openai/gpt-image-2.5/endpoints. Completed generations are billed in full; failed or cancelled ones are not billed.
Why not assume the cache discount?
Because Sume's page does not promise it. If a cached rate matters to your volume, treat it as unconfirmed and verify the billed usage.cost on real jobs.
How do I check this myself?
To find your real cost, run a few representative jobs and compare usage.cost on each response with the estimate you computed from the token rates. The linked docs pages and the catalog endpoint show the current values, and this post reflects them as of 2026-09-30.
Sources
Related posts
More in Models
- How many reference images? 10 on Image 1.0, 16 on GPT Image 2.5
Image 1.0 image_urls takes 1 to 10 public HTTPS URLs; ChatGPT Image 2.5 takes up to 16 references. Text-only models reject references.
- Image quality default on Sume: low on Image 1.0, high on GPT 2.5
Image 1.0 defaults quality to low; ChatGPT Image 2.5 on POST /v1/images defaults to high when omitted. Values, and when to raise quality.
- gpt-image-2 or openai/gpt-image-2? Model ids on Sume's Image API
Sume accepts bare Image Router ids such as gpt-image-2 and nano-banana-2 as aliases for openai/gpt-image-2 style ids. The catalog lists org/slug ids.
- Lyria 3.5 takes 10 images: Sume music takes one image_url
Google's Lyria 3.5 page lists up to 10 images alongside text. Sume's music request has a single optional image_url, so pick one accepted still per track.
Written by Sume