---
title: "OpenAI Releases List of Work Tasks It Says ChatGPT Can Already Replace"
description: "OpenAI has released a new evaluation to figure out how well its AIs perform on \"economically valuable, real-world tasks.\""
date: "2025-09-30"
modified: "2025-09-30"
authors:
  - name: "Victor Tangermann"
    job_title: "Senior Editor"
    link: "https://futurism.com/authors/victor"
url: "https://futurism.com/future-society/openai-work-tasks-chatgpt-can-already-replace"
categories:
  - "Artificial Intelligence"
  - "Future Society"
  - "OpenAI"
---

# OpenAI Releases List of Work Tasks It Says ChatGPT Can Already Replace

![OpenAI has released a new evaluation to figure out how well its AIs perform on "economically valuable, real-world tasks."](<https://futurism.com/wp-content/uploads/2025/09/openai-work-tasks-chatgpt-can-already-replace.jpg>)
*Getty / Futurism*

ChatGPT maker OpenAI has released a [new evaluation](<https://cdn.openai.com/pdf/d5eb7428-c4e9-4a33-bd86-86dd4bcf12ce/GDPval.pdf>), dubbed GDPval, to measure how well its AIs perform on "economically valuable, real-world tasks across 44 occupations."

"People often speculate about AI’s broader impact on society, but the clearest way to understand its potential is by looking at what models are already capable of doing," the company wrote in an accompanying [blog post](<https://openai.com/index/gdpval/>).

"Evaluations like GDPval help ground conversations about future AI improvements in evidence rather than guesswork, and can help us track model improvement over time," OpenAI added.

It's one of the most straightforward attempts to justify its AI models' financial viability to date, following skepticism that the tech may [prove to be a dead end](<https://futurism.com/ai-researchers-tech-industry-dead-end>). Experts have often [criticized the company's boastful marketing](<https://futurism.com/ceo-deepmind-openai-phd-ai>), such as CEO Sam Altman claiming that its GPT-5 model had [achieved "PhD-level" intelligence](<https://futurism.com/the-byte/rumors-openai-phd-human-intelligence>).

In "early results," GDPval found that "today’s best frontier models are already approaching the quality of work produced by industry experts" — a clear shot across the bow at critics who say the tech isn't up to the demands of the workplace.

The 44 occupations where "AI could have the highest impact on real-world productivity" included a litany of professions including real estate sales agents, social workers, industrial engineers, software developers, lawyers, registered nurses, customer service representatives, pharmacists, private detectives, and financial advisors.

The specific tasks, as laid out in a [paper](<https://cdn.openai.com/pdf/d5eb7428-c4e9-4a33-bd86-86dd4bcf12ce/GDPval.pdf>), range from creating a "competitor landscape for last mile delivery" for a financial analyst, assessing "skin lesion images" for a registered nurse, and designing a sales brochure for a real estate agent.

Surprisingly, the company found that its competitor Anthropic's Claude Opus 4.1 was the "best performing model" after being graded by industry experts across 220 tasks, followed by GPT-5, which "excelled in particular on accuracy."

An extra powerful version of GPT-5, called GPT-5-high, was "rated as better than or on par with the deliverables from industry experts" just over 40 percent of the time. GPT-4o, which was released more than a year ago, scored a mere 13.7 percent.

To be clear, OpenAI is treading carefully around the subject of replacing human jobs altogether. Its language suggests that AI will "support people in the work they do every day" instead of saying outright that anyone could soon be out of work because of AI. That's unsurprising, considering the negative optics of celebrating the loss of employment.

At the same time, whether that's really an honest interpretation of the industry's motives and end goals remains dubious. AI executives have long [boasted](<https://futurism.com/ceo-replacing-workers-ai>) about replacing human labor with AI — drastic cost-cutting measures that are [already starting to backfire for some companies](<https://futurism.com/companies-replaced-workers-ai>).

There's also good reason to take OpenAI's latest evaluation results with a massive grain of salt. We've already seen the use of AI cause major headaches for [software developers](<https://futurism.com/artificial-intelligence/new-findings-ai-coding-overhyped>), [lawyers](<https://futurism.com/judge-humiliating-punishment-lawyers-using-ai>), and even [customer service representatives](<https://futurism.com/klarna-openai-humans-ai-back>), often requiring more human oversight, not less.

Hallucinations, in particular, remain a major sticking point, undercutting the output of large language model-based tools, forcing users to [spend more time combing over the output](<https://futurism.com/ai-coding-programmers-reality>) of AIs for false information.

And while AI often excels at generating bursts of text in a particular style, it's easy for it to go off the rails during longer and less predictable tasks.

Real-world tasks are rarely "clearly defined with a prompt and reference files," OpenAI admitted.

"Early GDPval results show that models can already take on some repetitive, well-specified tasks faster and at lower cost than experts," the company wrote. "However, most jobs are more than just a collection of tasks that can be written down."

**More on OpenAI:** [*NBA Coach JJ Redick Says He Spends Hours Talking to His "Friend" ChatGPT*](<https://futurism.com/artificial-intelligence/jj-redick-nba-chatgpt-lakers>)

## Author
I've been at Futurism since 2017, where my role has evolved to encompass design, writing, and increasingly editing. I've always been fascinated by space exploration and advanced transportation, which I've leaned into by interviewing luminaries in those fields while closely following the dimensions of policy and regulation that allow next-generation projects to succeed -- or, sometimes, to fail. I'm also keenly interested in the effects of generative AI on society, policies, and democratic institutions, as well as clean energy, physics and biology, and the vagaries of tech leadership. My work for Futurism has been cited by publications including Ars Technica, Gizmodo, PC Magazine, Jalopnik, Fox News, and the New York Post. I spent my childhood living in locations including Manila, the Philippines, and Geneva, Switzerland, attended McGill University, and now live in Toronto, Canada. Before Futurism I worked at AskMen and a small photography studio. In my free time, I'm an avid gardener, foodie, and craft beer lover, as well as a maker of artisanal hot pepper sauces. I have a magnificent dog named Freida.

### Author social links  
[Bluesky](<https://bsky.app/profile/vtanger.bsky.social>)