AI news story

ChatGPT’s new Images 2.0 model is surprisingly good at generating text

ChatGPT Images 2.0, the newest image-generation model from OpenAI, shows just how much AI capabilities have evolved over the…

  • LLMs
  • Source: TechCrunch
  • Published: 2026-04-21

Editor's take

OpenAI’s latest image generation model, Images 2.0, integrated into ChatGPT, demonstrates an unexpected proficiency in rendering legible text within its visual outputs.

This development is significant because it addresses a long-standing limitation in generative AI for visual media, which has historically struggled with accurate typography and coherent text. For designers, marketers, and content creators who rely on AI for visual assets, this improvement could streamline workflows by reducing the need for post-generation text editing, a common bottleneck with earlier models like DALL-E 2. It signals a maturation of multimodal AI, moving beyond purely aesthetic generation to functional integration.

Future developments to monitor include the consistency of this text generation capability across diverse fonts, languages, and complex layouts, as well as OpenAI's approach to controlling for potential misuse of this feature, such as generating counterfeit documents or misleading signage. The performance of Images 2.0 relative to specialized text-in-image models will also be a key indicator of its market impact.