AI handwritten note in an image: write it, then read it back
A handwritten note in an AI image is still raster text. Generate it on GPT Image 2.5 at high quality, read it back, and fix errors with a masked edit.

To get a handwritten note inside an AI image, put the exact words in quotes in the prompt, ask for a handwriting style, and generate on a text-capable model such as openai/gpt-image-2.5 at quality: high. Then read the result back before you use it. Handwriting is drawn pixels, so a wrong letter needs a new edit, not a text change.
Prompt the words and the hand
Keep the note short. Quote the exact text, name the surface (lined paper, a sticky note, the back of a receipt), the pen (blue ballpoint, black marker) and the framing. Say that the writing must be legible and must not add other words. Short lines of five to ten words are easier to hold than a paragraph.
curl -X POST https://api.sume.com/v1/images \
-H "Authorization: Bearer $SUME_API_KEY" \
-H "Content-Type: application/json" \
-d '{"model":"openai/gpt-image-2.5","prompt":"Top-down photo of a yellow sticky note on a desk. Handwritten in blue ballpoint pen, legible, exactly: \"Back at 3. Coffee is on me.\" No other words.","quality":"high","aspect_ratio":"1:1"}'Pick the quality on purpose
Sume's GPT Image 2.5 rows accept quality: auto|low|medium|high|xhigh|max and default to high when you omit it. Dense or small writing is where a higher setting earns its cost. See the GPT Image 2.5 API guide for the sizes and limits.
Read it back
Treat every note as unverified until a person or an OCR step has compared it with the string you asked for. Look for swapped letters, doubled letters and invented words, and check the last line first because that is where drift tends to show. If the note is for a printed product, compare it at the real print size.
Fix one word, keep the rest
When one word is wrong, use a masked edit rather than a full re-roll. GPT Image 2.5 takes mask_url and up to 16 references. Mask the line, quote the corrected text, and say to keep everything outside the mask identical. The pattern is in mask_url edits on GPT Image 2.5.
When the words matter more than the look
If the exact wording is legal, financial or medical, do not rely on a generated hand. Generate the paper and pen texture, then set the words with a handwriting font in your layout tool, where the text is data you can proofread.
Settings that matter for text
These are the fields on the GPT Image 2.5 rows that affect a text-heavy image.
| Field | Values on openai/gpt-image-2.5 | Use for notes |
|---|---|---|
| quality | auto, low, medium, high, xhigh, max (default high) | high or above for small writing |
| aspect_ratio | auto, 1:1, 16:9, 9:16, 4:3, 3:4, 5:4, 9:8, 4:5 | 1:1 or 4:5 for a close-up note |
| n | 1 to 4 | ask for 4 and pick the cleanest |
| mask_url | public HTTPS mask | fix one line later |
Ask for more than one take
Set n to 4 and choose the best-written result. The endpoint price is per image, so four takes cost four times one take. For a short note this is usually cheaper than a long chain of corrections.
Related posts
More in Models
- AI holiday wreath product photos: Flux 2 Pro at $0.0375 each
Twelve wreath styles on Flux 2 Pro cost $0.45: $0.03 list times 1.25 is $0.0375 per image, with a 3:4 ratio and one reference photo per edit.
- AI ice cream flavor photos on Imagen 4 Fast: 40 images for $1.00
Imagen 4 Fast lists at $0.02 and bills $0.025 on Sume; ten flavors at four takes each is 40 images and $1.00, in 1:1, 4:3 or 3:4.
- AI meme template images on Grok Image: 13 ratios, $0.025 each
Grok Image lists at $0.02 and bills $0.025 on Sume, with 13 ratios from 2:1 to 9:20 and no 4:5; ten template backgrounds cost $0.25.
- AI photo booth strip from one selfie: a 1:4 Nano Banana 2 edit
A four-frame photo strip fits the 1:4 ratio that Nano Banana 2 lists on Sume; one selfie reference and one call cost $0.10 billed per strip.
Written by Sume