GPT Image, the current image generation model inside ChatGPT, represents enough of a shift from the older DALL-E models that habits built around prompting DALL-E can actually work against you now. The underlying approach changed from a purely from-scratch generation model to one that handles instructions and edits more conversationally, closer to how Nano Banana behaves.
The most noticeable practical improvement is in-image text rendering — DALL-E was notoriously unreliable at rendering legible text inside a generated image, while GPT Image handles this far more consistently, making poster-style, birthday-card, or caption-inside-image concepts genuinely usable now in a way they weren't reliably before.
Instruction-following has also improved meaningfully for complex, multi-part prompts. Where DALL-E often dropped or ignored some of the details in a longer prompt, GPT Image tends to respect a higher proportion of the specific instructions given, particularly around composition and object placement.
What to change in your prompting habits
Trust the model with more complex, multi-part instructions than you would have with DALL-E, and feel comfortable relying on it for concepts involving legible text, which was previously a weak point worth avoiding entirely.
What hasn't changed
It's still not the strongest choice for identity-preserving edits from a personal reference photo compared to Gemini's Nano Banana — GPT Image's strength is generation and text rendering, not photo-editing-style identity preservation specifically.
A practical takeaway
If you dismissed ChatGPT's image generation based on an older experience with DALL-E, it's worth revisiting — particularly for any concept involving readable text inside the image, where the improvement is the most noticeable.
Ready to create something?
Browse detailed AI image prompts and find a starting point for your next creative idea.
Explore AI Prompts →