---
title: "OpenAI’s New Image Generator Can Do Near-Perfect Text"
description: "OpenAI is rolling out a new image generator powered by its flagship GPT-4o model, and it nails rendering text."
date: "2025-03-26"
modified: "2025-03-26"
authors:
  - name: "Frank Landymore"
    job_title: "Contributing Writer"
    link: "https://futurism.com/authors/flandymore"
url: "https://futurism.com/openai-new-image-generator-perfect-text"
categories:
  - "Artificial Intelligence"
  - "OpenAI"
tags:
  - "chatgpt"
  - "image generator"
  - "OpenAI"
  - "sora"
---

# OpenAI’s New Image Generator Can Do Near-Perfect Text

![OpenAI is rolling out a new image generator powered by its flagship GPT-4o model, and it nails rendering text.](<https://futurism.com/wp-content/uploads/2025/03/openai-new-image-generator-perfect-text.png>)
*\<em\>Image: OpenAI\</em\>*

OpenAI is rolling out brand new image generation capabilities for ChatGPT. And guess what? It finally — *almost* — nails text.

Until now, the chatbot used the company's separate DALL-E model to dream up pictures. With this latest update, users will be able to access a new feature dubbed "Images in ChatGPT," leveraging OpenAI's flagship GPT-4o model, which has underpinned the chatbot for nearly a year. The upgrade is also available in Sora, OpenAI's video generation tool.

"This model is a step change above previous models," research lead Gabriel Goh [told *The Verge*](<https://www.theverge.com/openai/635118/chatgpt-sora-ai-image-generation-chatgpt>).

The most noticeable change is how the model handles text, something that it and its competitors have long struggled with. Words tended to come out looking like gobbledygook, and the text that was legible looked sloppy**,** filled with formatting errors and misspellings.

Not anymore, [according to OpenAI](<https://openai.com/index/introducing-4o-image-generation/>). One example shared by the company shows an employee writing out the pros and cons of the ChatGPT image update [on a whiteboard](<https://images.ctfassets.net/kftzwdyauwt9/5msykBd6Wu5mBcTgoqeJkj/4481c11698ff69f3d44d4c6220fade12/hero_image_1-whiteboard1.png>), following to the letter what was specified in the prompt; ditto for a [four-panel comic strip](<https://images.ctfassets.net/kftzwdyauwt9/6qMF89Gh1WqOVGrRSnzEIU/4e9013e2a0286bcdcde6d0160e39d5d8/ChatGPT_Image_Mar_24__2025__08_49_15_AM.png>) about a snail — all with cleanly rendered text.

https://twitter.com/OpenAI/status/1904602845221187829

"This was just like a process of iteration that took many, many months to get right," Goh told *The Verge*. "It's been just many months of small improvements." The model still struggles with [very small lettering](<https://images.ctfassets.net/kftzwdyauwt9/4MDiQPSrpvEnXlPLiIjkof/b2431ac47cfda64f59843d5570dc02d0/Screenshot_2025-03-24_at_3.59.45_PM.png>), but overall, the text quality is consistently usable, Goh said.

Unlike image generators like DALL-E, which use a diffusion model, GPT-4o uses an autoregressive approach that produces images from left to right and top to bottom, per *The Verge,* similar to how text — at least in English — is written.

Beyond improved penmanship, OpenAI says the model will now follow instructions better, as a common issue with older iterations was that they'd ignore certain details in lengthier prompts. It's also been fine-tuned to be able to generate [more photorealistic images](<https://images.ctfassets.net/kftzwdyauwt9/6qJSocGptmCSkDmsLRzIoD/a27270abadfc9f19d885cb752e9d6f74/boba.png>).

There are caveats. For one, it'll take longer to generate the outputs. And like all generative models, it's still prone to making up information, or hallucinating. It also struggles with generating non-Latin scripts, hallucinating characters when trying to write out languages like Korean.

With greater capabilities come greater safety and misinformation concerns. To this end, OpenAI stressed that it has particularly "robust safeguards" in place around nudity, violence, and [depictions of real people](<https://futurism.com/ai-taylor-swift-porn-consequences>). Moreover, all images that the AI model generates will be embedded with C2PA metadata identifying that it was made with GPT-4o. But this hidden watermark of sorts can easily be stripped — in fact, many social media platforms automatically remove an image's metadata once it's uploaded.

"Ultimately, no system is perfect for this type of thing, but we're continuously improving our safeguards and we think of this as a starting point," ChatGPT multimodal product lead Jackie Shannon told *The Verge*.

For now, GPT-4o image generation is only available to subscribers of OpenAI's ludicrous $200 per month Pro subscription tier, with plans to roll out the feature to Plus and free users in the near future.

**More on OpenAI:** *[Something Bizarre Is Happening to People Who Use ChatGPT a Lot](<https://futurism.com/the-byte/chatgpt-dependence-addiction>)*

## Author
At Futurism, my work has often centered on bringing a sense of clarity and insight to complex topics ranging from the regulation of emerging technologies to the esoteric ideologies of Silicon Valley executives, while striving not to lose the poetic sense of awe inspired by often-obscure fields like astrophysics and quantum computing. I broke the story of CNET using AI to produce articles that turned out to be riddled with factual errors and plagiarism — a dam-breaking inflection point, as I've reported, that's inspired copycats and endless discourse while beguiling stakeholders ranging from tech giants to purveyors of spam around the web. My work at Futurism has been cited by publications including CBS News, the Los Angeles Times, Vice, Gizmodo, Engadget, the Verge, and Vanity Fair. I grew up in locales ranging from India to China, and now live in the exotic suburbs of Virginia. In my free time, I'm an avid reader of weird sci-fi literature, an aficionado of East Asian cinema, and, regrettably, a relapsed gamer. Allegedly, I’m working on a debut novel, currently untitled.

### Author social links  
[Bluesky](<https://bsky.app/profile/f-w-l.bsky.social>)