AI meme image generator API: make the picture, then set the caption
For a meme, generate the image without words and add the caption in code, or quote it for GPT Image 2.5 and check the spelling. A Sume request.

To generate a meme with an image API, you have two options: have the model draw the caption into the picture, or generate a clean image and add the text yourself. Adding text yourself is more reliable. If you want the model to draw it, use a model that renders text well, put the caption in quotation marks, and read the result letter by letter. On Sume, openai/gpt-image-2.5 and Ideogram V3 are the first models to try for text (text rendering models).
Option A: text added in code
Ask for the scene with an empty band at the top or bottom: "a cat staring at a laptop, large empty space at the top, no text". Then draw the caption with Pillow or your editor. You can fix a typo without paying for a new image, and the same picture can carry captions in several languages.
Option B: the model draws it
Write the caption exactly, in quotes, and say how many times it appears and where: "The text 'WHEN THE BUILD PASSES' appears once, at the top, in white block capitals." A model can still double a word or drop a letter, so proofread the output (headline appears twice).
curl -X POST https://api.sume.com/v1/images \
-H "Authorization: Bearer $SUME_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "openai/gpt-image-2.5",
"prompt": "A cat staring at a laptop. The text 'WHEN THE BUILD PASSES' appears once at the top in white block capitals.",
"aspect_ratio": "1:1",
"quality": "medium"
}'Which to use
Pick by what you do with the output.
Mind rights: do not generate a real person's likeness or a trademarked character for a public post without clearance.
| Approach | Typo fix cost | Languages | Best for |
|---|---|---|---|
| Caption added in code | Free | Easy | Batches and localisation |
| Model draws caption | New generation | One per call | One-off posts |
Sources
Related posts
More in Use cases
- Ad music cutdowns: 15, 30 and 60 seconds from one AI track
Make 15, 30 and 60 second versions of one AI music bed: three prompts or one cut with a fade-out. Costs and limits from Sume's Music Router and Timeline docs.
- AI music for a tribute slideshow: a gentle bed for a memorial video
Score a tribute or memorial slideshow with AI music: an instrumental prompt, stills as timed holds, fades, and a listen-through before the video is shared.
- AI newsletter header image: generate a 4:1 banner and crop to fit
Newsletter headers are wide strips. Ask Nano Banana 2 for 4:1 or 8:1 on Sume, keep the subject in the middle, and crop to your template. Request and crop math.
- AI product angles from one photo: front, three-quarter and back views
Make extra product angles from one packshot: send a second photo of the back for printed sides, then check each view. A reference-edit request on Sume.
Written by Sume