---
title: "Anthropic CEO Admits We Have No Idea How AI Works"
description: "The CEO of Anthropic, one of the world's leading AI labs, just said the quiet part out loud — that nobody really knows how AI works."
date: "2025-05-04"
modified: "2025-05-04"
authors:
  - name: "Noor Al-Sibai"
    job_title: "Senior Staff Writer"
    link: "https://futurism.com/authors/nooralsibai"
url: "https://futurism.com/anthropic-ceo-admits-ai-ignorance"
categories:
  - "Anthropic"
  - "Artificial Intelligence"
tags:
  - "ai safety"
  - "anthropic"
  - "dario amodei"
  - "OpenAI"
---

# Anthropic CEO Admits We Have No Idea How AI Works

![The CEO of Anthropic, one of the world's leading AI labs, just said the quiet part out loud — that nobody really knows how AI works.](<https://futurism.com/wp-content/uploads/2025/05/anthropic-ceo-admits-ai-ignorance.jpg>)
*\<em\>Image: Saul Loeb / AFP via Getty / Futurism\</em\>*

The CEO of one of the world's leading artificial intelligence labs just said the quiet part out loud: that nobody really knows how AI works.

In an [essay published to his personal website](<https://www.darioamodei.com/post/the-urgency-of-interpretability>), Anthropic CEO Dario Amodei announced plans to create a robust "MRI on AI" within the next decade. The goal is not only to figure out what makes the technology tick, but also to head off any unforeseen dangers associated with what he says remains its currently enigmatic nature.

"When a generative AI system does something, like summarize a financial document, we have no idea, at a specific or precise level, why it makes the choices it does — why it chooses certain words over others, or why it occasionally makes a mistake despite usually being accurate," the Anthropic CEO admitted.

On its face, it's surprising to folks outside of AI world to learn that the people building these ever-advancing technologies "do not understand how our own AI creations work," he continued — and anyone alarmed by that ignorance is "right to be concerned."

But on another level, maybe it isn't; all the image and text generators that have exploded in popularity over the last few years work under the same principle of feeding in a gigantic pile of data and letting statistical systems mine it for patterns that can be reproduced. The whole thing is driven by ingested human creative works, not from first principles of machine intelligence.

"This lack of understanding," Amodei wrote, "is essentially unprecedented in the history of technology."

In Amodei's telling, that ignorance about how AI works and what unforeseen risks it may pose is a driving factor behind Anthropic.

In late 2020, the CEO and his sister Daniela left OpenAI amid concerns about the Sam Altman-run company's [safety practices](<https://www.nytimes.com/2023/07/11/technology/anthropic-ai-claude-chatbot.html>) — and in particular, that it was casting aside those concerns in pursuit of profit. The Amoideis and five other ex-OpenAI-ers [founded Anthropic](<https://techcrunch.com/2021/05/28/anthropic-is-the-new-ai-research-outfit-from-openais-dario-amodei-and-it-has-124m-to-burn/>) the next year to work on building safer AI — and part of that work seems to have been focused on figuring out the technology's nuts and bolts.

In recent months, Amodei wrote, Anthropic has begun to focus not only on helping "steer" AI — and its possible forthcoming progeny, artificial general intelligence — in ways that would benefit humanity, but also with the "tantalizing possibility" that researchers can finally figure out AI interpretability, or the "inner workings" of these systems, "before models reach an overwhelming level of power."

"Recently, \[Anthropic\] did an experiment where we had a 'red team' deliberately introduce an alignment issue into a model (say, a tendency for the model to exploit a loophole in a task) and gave various 'blue teams' the task of figuring out what was wrong with it," the CEO explained. "Multiple blue teams succeeded; of particular relevance here, some of them productively applied interpretability tools during the investigation."

While there will be a lot more work to be done to scale these "tools," which weren't explicitly detailed in Amodei's essay, it's still fascinating that folks at OpenAI's biggest competitor are not only working on making AI more advanced, but also tasking themselves with figuring out why and how it works.

"Powerful AI will shape humanity’s destiny," Amodei concluded, "and we deserve to understand our own creations before they radically transform our economy, our lives, and our future."

**More on Amodei:** [*Anthropic CEO Suggests That AI Deserves Workers' Rights*](<https://futurism.com/anthropic-ceo-suggests-ai-deserves-workers-rights>)

## Author
At Futurism, I've often been drawn to unpacking the narratives that underlie technological, scientific and medical progress, with a special interest in areas of conflict and ambiguity that end up setting agendas and steering the fates of both elites and the hoi polloi. I'm a committed generalist, but I often find myself returning to work involving NASA and the private space sector, the effects of AI on media and society, and the mechanics of the pharmaceutical industry, with a specific focus on the spread of GLP-1 drugs like Ozempic and Wegovy. Prior to Futurism, I worked for publications ranging from Media Matters and Truthdig to Raw Story and Bustle. I'm also the author of "Myspace Scene Queens," a 2024 title in Instar Books' acclaimed "Remember the Internet" series. My work at Futurism has been cited by outlets including the New Yorker, Slate, Nieman Lab, the Verge, the MIT Technology Review, the Sunday Times, and the Daily Beast. I grew up in North Carolina, attended the University of North Carolina at Asheville, and now live in Brooklyn, New York. In my free time, I'm an avid reader and music fan; you can probably find me at a local poetry reading, concert, underground rave, or DJ set. I'm the proud parent of an ineffable orange cat named Mee-Mow.

### Author social links  
[Bluesky](<https://bsky.app/profile/noorfromfuturism.bsky.social>)