Which AI image models cost under 4 cents on the Sume image API?
Six Sume image models list under $0.04 per image: Soul, Grok Imagine, Qwen Image, Imagen 4 Fast, Seedream 4.0 and Flux 2 Pro. What each takes, for drafts.

Six rows in the Sume image catalog list under four cents per image, read 2026-10-03: Soul at $0.005, Grok Imagine, Qwen Image and Imagen 4 Fast at $0.025 each, Seedream 4.0 at $0.0325 and Flux 2 Pro at $0.0375. Together they cover drafts, thumbnails and bulk variants where you plan to re-render only the winners on a more expensive model.
Prices are the output_image pricing line from GET /v1/images/models/{id}/endpoints; the Image API page says that line is what your wallet is charged and that failed generations are not billed.
The cheap rows and what each accepts
Flux 2 Pro is the sixth row and edits with up to 10 references at $0.0375. The other five are in the table.
| Model | Catalog id | Per image | References | `n` |
|---|---|---|---|---|
| Soul (Higgsfield) | higgsfield/soul | $0.005 | text only | 1 to 4 |
| Grok Imagine | x-ai/grok-image | $0.025 | edit up to 10 | 1 to 1 |
| Qwen Image | qwen/qwen-image | $0.025 | edit up to 10 | 1 to 4 |
| Imagen 4 Fast | google/imagen-4-fast | $0.025 | text only | 1 to 4 |
| Seedream 4.0 | bytedance-seed/seedream-4 | $0.0325 | edit up to 10 | 1 to 4 |
| Flux 2 Pro | black-forest-labs/flux.2-pro | $0.0375 | edit up to 10 | 1 to 4 |
Limits that matter for a draft pass
Each cheap row has a catch you should know before you build on it.
- Soul takes
num_imagesof 1 or 4 per the catalog description andresolutionof720por1080p; it is text-to-image only and lists nooutput_format. - Grok Imagine returns one image per call (
nrange 1 to 1), so a batch of four is four calls. - Imagen 4 Fast is text-to-image only and lists five aspect ratios (1:1, 16:9, 9:16, 4:3, 3:4).
- Qwen Image, Seedream 4.0 and Flux 2 Pro take references; the other three do not.
A draft-then-final budget
A concrete plan: 200 drafts on Qwen Image at $0.025 is $5.00. Pick the 20 best prompts and render each once on a pricier row, for example Ideogram V3 at $0.075 for $1.50. Total $6.50. Drafts only help when the cheap model's output predicts the expensive one's, which is a property you have to check on your own prompts; Sume does not state that it holds.
Remember that a call that exceeds the 30-second wait returns a 202 job instead of images, so a batch script should handle both status codes.
import os
import requests
resp = requests.post(
"https://api.sume.com/v1/images",
headers={"Authorization": f"Bearer {os.environ['SUME_API_KEY']}"},
json={
"model": "qwen/qwen-image",
"prompt": "flat-lay of a leather notebook and a pen",
"n": 4,
},
timeout=60,
)
resp.raise_for_status()
if resp.status_code == 202:
print("still running:", resp.json()["data"]["status_url"])
else:
for image in resp.json()["data"]:
print(image["url"])Sources
Related posts
More in Pricing
- AI lip sync API cost per second: H3 Max 480p, 768p, 1080p vs Fabric
On Sume, H3 Max lip-sync is $0.0625, $0.10 or $0.20 per audio second by resolution; VEED Fabric is $0.10 or $0.1875. A 14.8 s clip costs $0.94 to $3.00.
- AI music for 100 short videos: $12.50 flat on Sume Music
One hundred Sume Music generations cost $12.50 at the fixed $0.125 per accepted generation, whatever the prompt length. What that does and doesn't cover.
- AI video cost per deliverable: a rate sheet for finance approval
A cost-per-deliverable sheet built from Sume's published rates: 15-second avatar ad, captioned cut, 6-image set, voiced script, music bed, with a retake factor.
- AI video cost per minute by finished length, 15 s to 10 min
Finished AI video costs about $9.24 per minute at 15 s and $7.86 at 10 min with Wan 3.0, because captions and render round up. Full table, two models.
Written by Sume