12 thumbnail variants per video: cost on Sume by model
Twelve thumbnail variants per video cost $0.45 on Flux 2 Pro, $0.60 on Seedream 4.5 and $1.20 on Nano Banana 2 at 1K on Sume. Monthly totals for 30 videos.
Answer
Twelve thumbnail variants per video cost $0.45 on Flux 2 Pro, $0.60 on Seedream 4.5 and $1.20 on Nano Banana 2 at 1K, using Sume's catalog per-image prices on 2026-10-04. For 30 videos a month that is $13.50, $18.00 and $36.00.
Use one frame of the video, or a headshot, as an input_references entry, and write twelve short prompts that change only the background, the expression cue or the text area. Each variant is a separate generation, billed when it completes.
| Model | Per image | 12 variants | 30 videos |
|---|---|---|---|
| Flux 2 Pro | $0.0375 | $0.45 | $13.50 |
| Seedream 5.0 Lite | $0.04375 | $0.525 | $15.75 |
| Seedream 4.5 | $0.05 | $0.60 | $18.00 |
| Flux 2 Flex | $0.0625 | $0.75 | $22.50 |
| Nano Banana 2 (1K) | $0.10 | $1.20 | $36.00 |
Setup
- Pick a 16:9 ratio. Every row in the table lists 16:9, and Nano Banana 2 also lists 21:9 if you need an ultrawide banner.
- Send
n: 4three times per video on models whosenrange tops out at 4. That is three requests instead of twelve. - Keep text out of the render when you can. Ideogram 4.5 at $0.075 medium is the catalog's text-rendering row; the rows above are better at scenes than at long captions.
- Add the title text in your own editor or code after you pick the winner, so a typo never ships.
Cut the spend
Draft six variants on the cheapest edit row, pick two, and rerun only those two on a higher-priced model. Six drafts on Flux 2 Pro plus two finals on Nano Banana 2 costs $0.225 + $0.20 = $0.425 per video, a little less than twelve Flux 2 Pro images, with the better-looking pair at the end.
For larger batches, mode: "async" returns a job id at once and lets you poll, which avoids holding twelve connections open. See the Sume Image API docs for the request shape.
Before a large run
Prices and descriptors change when the catalog changes, so confirm them before you spend. Call GET /v1/images/models/{id}/endpoints for the row you plan to use and read its pricing line and supported_parameters; both come back in one response.
Then run a pilot of three to five images and read usage.cost on each response. Multiply by your planned count for a forecast you can trust. Completed generations are billed in full and failed or cancelled ones are not, so a pilot that errors costs nothing.
For big batches, use mode: "async" or mode: "webhook" with a public HTTPS webhook_url, so no request waits on the 30-second sync limit. Poll GET /v1/jobs/{id}/status and fetch GET /v1/jobs/{id}/result when the job completes.
Sources
Related posts
More in Use cases
- A 15-second Omni ad: two clips, one fade, what it costs
Omni makes 3 to 10 seconds, so a 15-second ad is a 10 and a 5 second clip joined by a Sume Timeline fade: about $1.98 at 720p versus $1.50 of Omni on Google.
- 200-image mood board for $1: Soul at half a cent on Sume
Higgsfield Soul costs $0.005 per image on Sume, returns 1 or 4 per call, and is text-only. 200 mood board images cost $1.00 in 50 calls of four.
- A 30-minute lofi study video from 12 Lyria tracks for $4.51
Twelve 2.5-minute Lyria tracks cost $1.50, joining them $0.01 and a 30-minute Timeline render $3.00. Plan it free with the unbilled plan call.
- 90-second AI video API: three 30 s Wan 3.0 jobs and what they cost
A 90-second AI video on Sume is three 30-second jobs. Wan 3.0 totals $5.64 at 480p, $11.25 at 720p and $22.50 at 1080p. A shot-list plan with a Python loop.
Written by Sume