Get Ready to be Amazed: OpenAI’s New Image Generator is Here

Get Ready to be Amazed: OpenAI’s New Image Generator is Here
  • calendar_today August 10, 2025
  • Technology

OpenAI now offers a new feature called “Images in ChatGPT” which allows users to produce images directly from the ChatGPT interface through seamless integration. The GPT-4o model powers this advancement which lets users generate images during chat while delivering a major progression in AI content creation.

OpenAI has expanded access to advanced image generation functionality by making it available to all ChatGPT subscription levels including Plus, Pro, Team and the free version. The current limits for free tier users who produce about three images per day match those of DALL-E 3 but OpenAI’s Taya Christianson stated these restrictions might change as demand fluctuates. The dedicated custom GPT will ensure that DALL-E enthusiasts maintain access to the platform.

OpenAI research lead Gabriel Goh emphasized GPT-4o’s transformative capability describing it as an “omnimodal” model which processes text, images, audio and video data. The model now features improved “binding” abilities which solve a longstanding issue in AI image generation. GPT-4o provides reliable management of 15 to 20 objects without experiencing color or shape confusion whereas earlier models misinterpreted such relationships.

The system’s advanced text rendering capabilities stand out as a significant development. The generation of AI images has historically produced text which is either distorted or lacks logical coherence. Goh explained the meticulous development process as a lengthy iterative procedure which required months to perfect properly. The team has reached a consistency level that makes text in images reliably usable despite the ongoing challenge of achieving perfect text rendering for small text.

The system uses an autoregressive design instead of the diffusion models that are standard for image generation. The sequential image generation technique from left to right and top to bottom mirrors text generation principles and enhances text rendering and binding capabilities.

OpenAI presented several uses of their system during a briefing session which demonstrated its ability to generate scientific diagrams with accurate labeling like Newton’s prism experiment as well as multi-panel comics with consistent characters and dialogue alongside informational posters with precise text. The system demonstrated practical uses including the creation of transparent background images for stickers and restaurant menus along with logos.

ChatGPT’s multimodal product lead Jackie Shannon highlighted the system’s capability to utilize vast world knowledge. She explained that her image drawing process combines her personal skill constraints with the comprehensive world knowledge she has acquired. The model incorporates world knowledge into its process which allows users to request images of Newton’s prism experiment without needing to provide any background information.

OpenAI explains that despite a slight increase in the time required for image generation, the improved quality and expanded capabilities make the wait worthwhile. Even though latency improvements remain a goal we strive for, Shannon emphasized that the superior quality of the images and their advanced capabilities, together with global knowledge, compensate for any additional waiting time.

Key Features and Safeguards Implemented by OpenAI:

  • Enhanced Binding: GPT-4o maintains precise connections between 15 and 20 objects, which helps reduce confusion between colors and shapes.
  • Improved Text Rendering: Careful development leads to dependable text rendering in generated images by solving a frequent AI problem.
  • Autoregressive Approach: The sequential method of image generation employed by the system offers potential improvements for handling text and objects.
  • Robust Safeguards: OpenAI protects against watermark removal and sexual deepfakes while denying requests to produce CSAM.
  • C2PA Metadata: Standard C2PA metadata is embedded within all images generated by the system to identify them as products of OpenAI.
  • User Ownership: Users maintain ownership rights over images they create as long as they adhere to the established usage policies.

OpenAI made clear its commitment to preventing potential misuse through the establishment of powerful protective measures. According to Shannon, no system exists that is flawless for this task, yet ongoing improvements to our safeguards represent our initial effort. All images produced through ChatGPT remain under user ownership and may be used freely in line with our established usage policies.

OpenAI enhances ChatGPT capabilities through “Images in ChatGPT,” while establishing a benchmark for accessible AI image creation and simultaneously mitigating associated technological risks.