- calendar_today August 10, 2025
OpenAI introduced “Images in ChatGPT,” which integrates image generation functionality directly into the ChatGPT interface as a revolutionary capability. The new feature enabled by the recently introduced GPT-4o model allows users to generate images during conversations, which represents a major advancement in AI-driven content creation.
The “Images in ChatGPT” option is now available to users on all ChatGPT subscription plans, including Plus, Pro, Team, and free accounts, to make advanced image creation more accessible. Users of the free tier can generate around three images daily, but OpenAI spokesperson Taya Christianson said these limits could change with fluctuations in demand. Users preferring a separate DALL-E experience will find access through a custom GPT interface.
OpenAI research lead Gabriel Goh described GPT-4o as an “omnimodal” system that can handle multiple data types such as text, images, audio, and video. The model now displays better “binding” functionality, which resolves one of AI image generation’s major obstacles. GPT-4o can accurately handle 15 to 20 objects without confusing their colors or shapes, addressing previous models’ difficulties in maintaining object-attribute relationships.
The system now features an advanced text rendering capability that stands out as a significant development. AI-generated images of the past frequently contained text that appeared scrambled and meaningless. According to Goh, the creation process was a long series of iterations that required many months to achieve proper functionality. Despite the ongoing challenge of perfect text rendering for small text, the team now provides consistent results that make image text reliably usable.
Rather than using typical diffusion models, which most image generators utilize, this system’s design implements an autoregressive approach. The generation method builds images from left to right and top to bottom, following a text-like process, which is believed to support better text rendering and binding abilities.
OpenAI demonstrated the system’s wide range of uses at a briefing, which included producing labeled scientific diagrams such as Newton’s prism experiment while maintaining character continuity and dialogue in multi-panel comics, as well as creating informational posters with precise text. The system demonstrated practical uses by producing transparent background images for stickers and restaurant menus, as well as logos.
Jackie Shannon highlighted how the system utilizes extensive world knowledge during product development. She explained that when she begins to draw an image, she works within her own skill boundaries, yet incorporates the world knowledge she has accumulated. With world knowledge integrated into the model, users can request an image of Newton’s prism experiment and receive it without needing to provide any additional information.
OpenAI maintains that the improved quality and expanded capabilities of their image generation process make the increased wait time worthwhile. Shannon acknowledged that although there’s potential to improve latency times, the exceptional quality and capability of these images, combined with their world knowledge, compensate for the extra waiting time.
OpenAI addressed potential misuse concerns by emphasizing its robust safeguard implementation. The system incorporates safeguards to defend against watermark removal while simultaneously blocking sexual deepfake production and refusing CSAM-related requests. All generated images will contain standard C2PA metadata to identify them as OpenAI creations even though they lack visual watermarks. The company operates its own image verification tools internally.
According to Shannon, no system achieves perfect protection, but OpenAI keeps enhancing its safeguards, which they see as an initial step. Users who generate images from ChatGPT have full ownership rights, enabling them to utilize these images according to our usage policies in any way they desire.
The addition of advanced image creation capabilities to ChatGPT marks a major advancement in AI-powered creative technology. The dedication to enhanced binding features combined with advanced text rendering and strong safeguard measures shows how OpenAI is delivering a tool that is both powerful and ethically sound. The company’s innovative image generation strategy stands out because it involves switching to an autoregressive approach instead of using traditional diffusion models. OpenAI underscores transparency and ethical usage through its focus on user ownership and metadata integration in the development of AI-generated content. The launch establishes a new benchmark for AI image generation systems that are both user-friendly and powerful, alongside addressing associated risks.





