OpenAI's image models, offered here in four tiers: GPT Image 2 at 2K and 4K, and the newer 2.5 at 2K and 4K. They are the models to pick when the picture has to contain words someone can read.
Sign in to start — new accounts get 20 free credits, no card.
The same prompts sit above the input box; click one there to load it.
Matte black ceramic mug on a pale linen surface, soft window light from the left, shallow depth of field, editorial product photography
Close-up portrait of an older woman laughing, natural window light, film grain, 85mm lens, shallow depth of field
Misty pine forest at dawn, low fog between the trunks, cold blue light, wide shot
Images generated on this site with the same models.



Signage, packaging, book covers, UI mockups. Put the exact words in quotes and these are the models most likely to render them as letters rather than as shapes that resemble letters.
The newer generation holds on to more of a complex prompt. If a detailed instruction keeps getting partly ignored on the cheaper tiers, this is the one to try.
The 4K tiers cost more because they produce more. Our measured output at the 2K tier was already 2048 × 2048 PNG.
Up to 16 input images, which makes these tiers usable in Virtual Try-On and any other multi-image workflow.
Pick any of these from the Model menu above; the price is shown before you generate. Each name links to what that model is good at. Compare every tier.
Google Gemini 2.5 Flash · fastest, cheapest
5 credits per image
Black Forest Labs · fastest, about 5s per image
5 credits per image
ByteDance · good at Chinese text and posters
7 credits per image
ByteDance · most photorealistic, highest detail
18 credits per image
OpenAI · good with text in images, half the price of 2.5
7 credits per image
OpenAI · for print or heavy cropping
10 credits per image
OpenAI · newest, better instruction following
13 credits per image
OpenAI · newest, highest resolution
20 credits per image
Seven credits. Enough to find out whether the model understands the brief.
2.5 for a long instruction it keeps dropping; 4K for print or heavy cropping. Otherwise the cheaper tier is the same picture.
Text is the reason to use these models and also the first thing to go wrong. Zoom in before you download.
"a sign reading 'OPEN DAILY 7-3'". Unquoted text gets paraphrased or invented.
A few words render reliably. A paragraph does not, on any model here.
"Centred on the awning", "along the bottom edge". Placement is part of the picture, not an afterthought.
4K buys pixels, not comprehension. If the model is misreading the brief, go to 2.5 rather than to 4K.
Start free. Every tool draws from one balance — no separate subscriptions.
$0 / month
Enough to see whether the results are any good
$6.9 / month
Enough to find out if this works for you
$13.9 / month
For people editing every week
$39 / month
For teams and heavy output
2.5 is the newer generation and holds on to more of a long instruction. Both are offered at 2K and 4K. Version 2 at 2K costs 7 credits; 2.5 at 2K costs 13.
GPT Image 2 · 2K at 7 credits. Move to 2.5 when the model keeps dropping part of your instruction, and to 4K only when you need the pixels for print or cropping.
Yes, alongside Seedream for poster layouts. Put the exact words in quotes and keep them short.
Only if you need the resolution. The picture is the same picture; you are paying for pixels. For a larger file from a cheaper tier, run the 2K output through the Image Upscaler for 2 credits.
You are only charged when an image comes back. A failed or timed-out run costs nothing.
GPT Image 2.5 · 2K is preselected above. Switch tiers in the model menu.
All of them are in the same dropdown, and one credit balance covers every one.