---
title: "GPT-5 Launch Demo Plagued With Catastrophically Dumb Errors"
description: "OpenAI's attempt to show off its latest GPT-5 model's awesome performance states produced wildly embarrassing gaffes."
date: "2025-08-08"
modified: "2025-08-08"
authors:
  - name: "Frank Landymore"
    job_title: "Contributing Writer"
    link: "https://futurism.com/authors/flandymore"
url: "https://futurism.com/gpt-5-demo-dumb-errors"
categories:
  - "Artificial Intelligence"
  - "OpenAI"
tags:
  - "artificial intelligence"
  - "chatgpt"
  - "gpt-5"
  - "OpenAI"
---

# GPT-5 Launch Demo Plagued With Catastrophically Dumb Errors

![OpenAI's attempt to show off its latest GPT-5 model's awesome performance states produced wildly embarrassing gaffes.](<https://futurism.com/wp-content/uploads/2025/08/gpt-5-demo-dumb-errors.jpg>)
*\<em\>Image: @EgeErdil2 via X / Futurism\</em\>*

OpenAI's GPT-5 is finally here and already powering ChatGPT, but it [hasn't made a great first impression](<https://futurism.com/gpt-5-sucks>).

In a livestream dedicated to the release, OpenAI tried to show off its newest large language model which CEO Sam Altman called a "significant step along the path to AGI"— but instead turned heads with some catastrophically dumb errors.

Across several examples, bar graphs intended to show off GPT-5's awesome performance benchmarks, while appearing professional-looking, turned out to be horribly inaccurate nonsense upon closer inspection.

The gaffes were [flagged on social media](<https://x.com/shreyk0/status/1953509438255464603>) and [highlighted by *The Verge*](<https://www.theverge.com/news/756444/openai-gpt-5-vibe-graphing-chart-crime>). The [most egregious example](<https://x.com/EgeErdil2/status/1953505551570415718>) is a bar graph comparing coding benchmark scores for GPT-5 compared to older models. Somehow, the bar for GPT-5's score of 52.8 percent accuracy is nearly twice as tall as the bar for a score of 69.1 percent for the o3 model. Even more bafflingly, the 69.1 percent bar is the exact same size as another bar representing 30.8 percent for GPT-4o. Make it make sense!

https://twitter.com/EgeErdil2/status/1953505551570415718

OpenAI hasn't confirmed if it used GPT-5 to generate the graphs — and at this point, it has every reason not to — but it's an incredibly embarrassing mistake from a company that's [being valued](<https://www.ft.com/content/ab1ef47e-3c5b-49e0-afea-7c8be9d351e4>) in the region of half a trillion smackeroos.

It's also a little poetic. Some research suggests that newer models could actually be [getting dumber](<https://futurism.com/ai-industry-problem-smarter-hallucinating>) in key ways, hallucinating more frequently than earlier versions. One [study](<https://venturebeat.com/ai/anthropic-researchers-discover-the-weird-ai-problem-why-thinking-longer-makes-models-dumber/>) even found that the longer these new reasoning models "think," the more their performance deteriorates. Other research implicates the AI slop that's [increasingly poisoning the AI's training data](<https://futurism.com/ai-models-falling-apart>). Circling back to GPT-5's bar graph, you have OpenAI trying to spin its lower score of 52.8 as actually being better than its predecessor's.

Altman, playing it cool, tried to laugh off the blunder.

"\[W\]ow a mega chart screwup from us earlier," he [tweeted](<https://x.com/sama/status/1953513280594751495>), in his typical lower-case patois. "wen GPT-6?!"

OpenAI corrected the charts in its [blog post](<https://openai.com/index/introducing-gpt-5/>), but the originals are still there in the livestream.

Human error may or may not be to blame for the charts, but following GPT-5's release, users were quick to expose how error-prone its image- and diagram-generating capabilities remain. One asked ChatGPT to draw a map of two cities in Virginia with their neighborhoods labeled, prompting it to return names that were [complete gobbledygook](<https://bsky.app/profile/pulpandpolitics.bsky.social/post/3lvvgdaxz522m>).

And in what should've been a layup for GPT-5, Ed Zitron of the "[Where's Your Ed At?](<https://www.wheresyoured.at>)" newsletter [found](<https://bsky.app/profile/edzitron.com/post/3lvua4fgc722k>) that the AI couldn't even nail a simple map of the US. Ever think of visiting "West Wigina," "Delsware," "Fiorata," or "Rhoder land"? Or maybe "Tonnessee" and "Mississipo?"

The irony is that OpenAI [bragged](<https://futurism.com/openai-new-image-generator-perfect-text>) back in March that an update for its previous GPT-4o model meant that ChatGPT could now excel at generating texts in images.

"As you can tell now it's very good at text," one of the [example generated images](<https://x.com/OpenAI/status/1904602845221187829>) read. "Look at all this accurate text!"

Sounds like they might've spoken too soon. Or maybe AI models really are going backwards.

**More on OpenAI:** *[GPT-5 Users Say It Seriously Sucks](<https://futurism.com/gpt-5-sucks>)*

## Author
At Futurism, my work has often centered on bringing a sense of clarity and insight to complex topics ranging from the regulation of emerging technologies to the esoteric ideologies of Silicon Valley executives, while striving not to lose the poetic sense of awe inspired by often-obscure fields like astrophysics and quantum computing. I broke the story of CNET using AI to produce articles that turned out to be riddled with factual errors and plagiarism — a dam-breaking inflection point, as I've reported, that's inspired copycats and endless discourse while beguiling stakeholders ranging from tech giants to purveyors of spam around the web. My work at Futurism has been cited by publications including CBS News, the Los Angeles Times, Vice, Gizmodo, Engadget, the Verge, and Vanity Fair. I grew up in locales ranging from India to China, and now live in the exotic suburbs of Virginia. In my free time, I'm an avid reader of weird sci-fi literature, an aficionado of East Asian cinema, and, regrettably, a relapsed gamer. Allegedly, I’m working on a debut novel, currently untitled.

### Author social links  
[Bluesky](<https://bsky.app/profile/f-w-l.bsky.social>)