
Tech • AI • Robotics
OpenAI has introduced GPT Image 2.5, adding in-chat image creation, sketch-based prompting, and stronger multi-step editing consistency, with early hands-on comparisons suggesting better visual quality than Gemini Nano Banana Pro but slower generation speed.
GPT Image 2.5 can be used in the web interface, desktop app, developer workflows, and through the API. Two API variants were highlighted: GPT Image 2.5 Flare for faster generation and GPT Image 2.5 Sunburst for more detailed creative work with longer render times. The release expands image generation beyond a single prompt box into a broader editing and asset-creation workflow.
The updated interface includes an Images section with templates and trending prompts, alongside standard chat-based generation. That signals a shift toward more structured visual creation, where users can start from preset formats or move directly into custom requests for assets, illustrations, and design elements.
A new sketch tool lets users draw directly with a mouse or finger, then turn rough shapes into finished images. In one test, a crude sunset drawing with hills and sky was converted first into a cartoon-like scene and then, after a follow-up request, into a more realistic landscape. The result points to stronger interpretation of low-quality visual input and smoother prompt-based refinement.
Images can now be edited through a visual toolset that includes commenting on selected areas, erasing elements, removing backgrounds, resizing, and markup. A simple test removed clouds from a circled area of a generated landscape while preserving the rest of the composition, showing a more direct and accessible inpainting workflow.
The central improvement is multi-turn editing consistency. Earlier image models often drifted away from the original subject or composition after several revisions, but GPT Image 2.5 is designed to keep prior changes intact across longer editing chains without visibly degrading quality. That matters for repeated asset iteration, especially in automated or agent-based production pipelines.
In a benchmark shared for the release, a sequence of 150 generated frames showed the difference between earlier and newer behavior. The older model produced noticeable movement drift in a cube animation, while GPT Image 2.5 held much closer to the intended pose guidance, resulting in smoother output. A second example with candles lighting one by one showed minimal unwanted change between frames.
In side-by-side tests against Gemini Nano Banana Pro, GPT Image 2.5 was judged stronger for thumbnail-style images. When asked to improve an existing thumbnail, it produced an output with more vivid color, stronger lighting, and a more polished overall look. Using the same prompt on both systems, the newer OpenAI model was also seen as producing more convincing skin texture and a less synthetic appearance, though some AI-generated cues remained.
A second comparison focused on a realistic infographic featuring landmarks and informational text. GPT Image 2.5 was reported to follow instructions more closely, producing a result with clearer realism and more complete monument details than the Nano Banana Pro version. In that test, the OpenAI output was rated roughly 8/10 versus 6/10 for the competing model.
Despite claims of faster performance relative to the earlier GPT Image model, real-world comparison suggested Nano Banana Pro still rendered faster in some side-by-side runs. The practical takeaway is that GPT Image 2.5 appears to improve quality and consistency, but users may need to accept slower turnaround on some generations.
GPT Image 2.5 appears to be a meaningful upgrade in editing reliability, prompt adherence, and visual polish rather than a complete break from previous image models. Its strongest value may be in workflows that depend on repeated revisions without losing subject consistency.
Explain this