The new ChatGPT Images feature marks a significant step in how OpenAI integrates visual generation into its conversational assistant. With GPT Image 1.5, the company strengthens ChatGPT's graphics capabilities with increased speed, improved editing, and greater control over the final result, in the midst of a competitive race with Google and its Gemini ecosystem.
The update goes beyond simply producing more attractive images; it aims to ensure the system fully respects the user's intent with each modification. The goal is for ChatGPT to evolve from an advanced text chat into a space where writing, editing, and reusing images feels much more like working in a complete creative studio.
What is ChatGPT Images and what does GPT Image 1.5 offer?
With ChatGPT Images, OpenAI directly integrates an image generation and editing model into the assistant's interface , capable of working from scratch or from existing photographs or designs. At the heart of this innovation is GPT Image 1.5 , a revamped version of the visual model already used in ChatGPT, now optimized to follow complex instructions with much greater accuracy.
The key difference lies in the model's ability to modify only what it's asked to : if the lighting, background, or a small detail of the scene is changed, the rest of the image remains consistent. Elements such as facial features, framing, overall style, and composition remain stable even after several rounds of changes—something that until now has been a major weakness of most AI generators.
This ability to maintain consistency is especially noticeable when working with user-uploaded images . The system understands much better which parts of the photo should be preserved and which should be transformed, avoiding the "completely new image" effect that so often forced the process to be restarted from scratch.
Furthermore, GPT Image 1.5 stands out for its more reliable tracking of long, sequential instructions . The model can apply changes step by step without losing sight of the original structure, enabling more complex compositions where the relationships between elements (distance, relative position, proportions) are maintained as requested.

Speed and workflow: up to four times faster imaging
Another key aspect of the release is the substantial improvement in response times. OpenAI claims that GPT Image 1.5 generates images up to four times faster than the previous model, reducing the waiting times that used to hinder creative work when running multiple tests in succession.
The idea is that users can test, correct, and refine designs almost continuously , without latency being a hindrance. While an image is being processed, new requests can be made or alternative approaches explored, which aligns better with real-world workflows in design studios, marketing agencies, or digital content departments.
The company also emphasizes that the new model is not only faster, but also produces more usable results on the first try . Rendering many small faces in a single scene, the natural appearance of materials or backgrounds, and the integration of added elements within the original lighting have all been areas of improvement, reducing the need to repeat the same request so many times.
From a productivity standpoint, this combination of increased speed and greater accuracy aims to transform ChatGPT Images from a simple experimental tool into a viable option for professional projects. Both independent creators and content teams can then dedicate more time to deciding what they want to share and less time struggling with the model.
Iterative editing, text on image, and creative control
One of the major changes in approach in GPT Image 1.5 is its iterative editing . Instead of generating a completely new image each time the user modifies a instruction, the model focuses on applying the requested adjustments to the existing database. This is crucial in projects where visual consistency —for example, in advertising campaigns, brand identities, or product catalogs—is as important as the final result.
The model performs particularly well in tasks such as adding, removing, combining, merging, or transposing elements while maintaining the key features that make a scene recognizable. You can ask it to add an object to the background, change a person's clothing, or transform the environment (from day to night, from summer to winter) without the model "forgetting" the overall structure of the image.
Added to this is a key improvement: the rendering of text within images . GPT Image 1.5 is capable of rendering denser text with smaller font sizes, opening the door to creating posters, menus, infographics, and editorial pieces where typography no longer appears distorted or unrecognizable. In a European context, where signage, informational documents, and corporate materials require legibility, this advancement is particularly relevant.
OpenAI notes that the model is also more “creative” in how it modifies and adds design elements . From relatively simple prompts, ChatGPT Images can propose reasonable and visually consistent compositions, even without extremely detailed prompts. This lowers the barrier to entry for users who are not used to writing complex instructions.
At the same time, for those who do want fine control, the system responds better to long strings of instructions , allowing for incremental adjustments to colors, styles, element layout, and fonts. The combination of both approaches—guided exploration and detailed control—aims to accommodate both expert creative professionals and more general users.
A dedicated space within ChatGPT: towards a visual "studio"
The launch of ChatGPT Images is not just a change in model, but also in user experience. OpenAI introduces a dedicated space for images in the ChatGPT interface , accessible from the sidebar, which functions more like a small creative studio integrating styles, filters, and suggestions.
In this environment, users can explore predefined styles , try filters, modify a pre-generated image, or work from a user-uploaded photograph, all without needing to write complicated instructions. It's a way to make visual generation more accessible to those who are more accustomed to mobile apps or online editors than to the lengthy, traditional AI prompts.
The company says the overall ChatGPT experience will incorporate more visual and multimodal elements, such as voice mode, in other areas of the assistant as well. Search results, explanations of concepts, and practical queries—like unit conversions or sports data—could increasingly rely on charts, diagrams, or flowcharts to make the information clearer.
The underlying philosophy is that when an answer is better explained with an image than with text alone, that visual support should be offered natively . This approach aligns with the trend in educational platforms, digital media, and productivity apps in Europe, where combining text and images has become standard practice for improving comprehension and retention.
In parallel, this visual space aims to reduce the distance between the initial idea and the final result: fewer jumps between tools, fewer back-and-forths between chat, image generator and external editors, and more ability to take an idea from the sketch to a near-final version without leaving the same environment.
Competition with Google Gemini and Nano Banana Pro
The arrival of ChatGPT Images and GPT Image 1.5 comes amid direct competition with Google in the field of generative AI. In recent months, the search engine has gained visibility with its Gemini family of models and, in particular, with tools like Nano Banana Pro (Gemini 3 Pro Image) , which boast greater world knowledge and a robust ability to render text in images.
These advances have allowed Google to take the lead in several specialized rankings, which has set off alarm bells at OpenAI. Internally, there has even been talk of a "code red" to describe the need to react quickly and regain ground in market share and technological perception, especially among developers and companies.
In this context, GPT Image 1.5 presents itself as a direct response, with a slightly different approach: Google focuses on ultra-fast generation and visual improvisation , useful for quick ideas or prototypes, while OpenAI emphasizes iterative editing and visual consistency. To put it simply, Nano Banana Pro tends to "reinterpret" the scene with each change, while ChatGPT Images attempts to build version by version without completely discarding previous work.
This difference in philosophy is especially relevant for design, marketing, media and e-commerce in Europe , where consistent images are needed over time: campaigns that maintain the same aesthetic, catalogs that preserve the product's appearance, or informational pieces that must conform to an established graphic style.
OpenAI's offensive is not an isolated event. It follows moves such as the launch of GPT-5.2 , aimed at developers and professionals, and is part of a broader strategy to strengthen its position in both the text and visual aspects of generative AI in the face of the Gemini ecosystem's momentum.
Availability, uses and current limitations of the model
OpenAI has begun the widespread rollout of GPT Image 1.5 in ChatGPT , enabling any user to start generating and editing images directly from the assistant's interface. Furthermore, the model is also accessible through the company's API , allowing European developers to integrate its capabilities into applications, websites, or internal tools.
For the Business and Enterprise versions, the company plans a gradual rollout , so that businesses can test the new features in their visual workflows, from creating marketing materials to generating internal graphic resources for documentation or training.
Although the leap in quality is evident, OpenAI acknowledges that the model still has significant limitations . These include the difficulty in maintaining the exact identity of each person when a large group appears in the same image, and certain inconsistencies in translations and text rendering for languages such as Chinese, Arabic, or Hebrew , where typography and writing direction pose additional challenges.
Despite this, the company emphasizes that GPT Image 1.5 significantly improves precise editing and the creation of complex compositions compared to previous generations. The model is especially geared towards those who need to iterate on the same image many times without losing the details that make it recognizable.
On a more playful level, ChatGPT Images also fuels the use of visual AI as an entertainment tool . The ability to reimagine a person as a historical figure, athlete, superhero, or protagonist in fantasy scenes remains a major draw for the general public, and now it can be done with greater speed and control.
With all these new features, OpenAI's proposal for images becomes more ambitious: a complete platform where you can plan, create and refine visual content continuously, reconciling creative experimentation with the demands of professional projects in Spain and the rest of Europe.