---
title: "Anthropic Was So Concerned About Its New Mythos-Based Model’s Power That It Lobotomized Its Ability to Improve Itself"
description: "Anthropic announced Mythos 5 and a slightly less powerful model, called Fable 5, which the company is limiting, citing security concerns."
date: "2026-06-11"
modified: "2026-06-11"
authors:
  - name: "Victor Tangermann"
    job_title: "Senior Editor"
    link: "https://futurism.com/authors/victor"
url: "https://futurism.com/artificial-intelligence/anthropic-concerned-models-ability-improve-itself"
categories:
  - "Anthropic"
  - "Artificial Intelligence"
  - "Blockchain"
  - "Cryptocurrency"
  - "Future Society"
---

# Anthropic Was So Concerned About Its New Mythos-Based Model’s Power That It Lobotomized Its Ability to Improve Itself

![A stylized photo illustration featuring Anthropic co-founder Dario Amodei.](<https://futurism.com/wp-content/uploads/2026/06/anthropic-concerned-models-ability-improve-itself.jpg>)
*Illustration by Tag Hartman-Simkins / Futurism. Source: Michael M. Santiago / Getty Images; Shutterstock*

Earlier this year, Anthropic refused to release its Mythos AI model to the public, [saying it was simply too dangerous](<https://futurism.com/artificial-intelligence/anthropic-claude-mythos-escaped-sandbox>).

At the time, executives claimed the model was capable of [punching through powerful cybersecurity safeguards](<https://futurism.com/artificial-intelligence/security-experts-alarmed-anthropic-mythos>), pointing at researchers who used it to discover [thousands of vulnerabilities](<https://www.reuters.com/technology/anthropic-rolls-out-public-version-mythos-without-cybersecurity-capability-2026-06-09/>) in widely-used open source code.

Months later, Anthropic was finally ready to go public with the model. On Tuesday, the Dario Amodei-led company [announced](<https://www.anthropic.com/news/claude-fable-5-mythos-5>) a Mythos-powered model called Fable 5, which it claims is "safe for general use."

However, new safeguards quickly frustrated AI researchers, who accused the company of intentionally lobotomizing Fable 5. The backlash was so fierce, Anthropic quickly made adjustments to the policy, as [*Wired* reported](<https://www.wired.com/story/anthropic-responds-to-backlash-on-claudes-secret-sabotage-on-ai-research/>) on Wednesday, highlighting just how carefully the company is treading.

In its original announcement, Anthropic [claimed](<https://www-cdn.anthropic.com/d00db56fa754a1b115b6dd7cb2e3c342ee809620.pdf>) the safeguards were designed to stop Fable 5 from improving itself, in "new interventions that limit Claude’s effectiveness for requests targeting frontier LLM development." Just days ahead of the launch, Anthropic [released a report](<https://www.anthropic.com/institute/recursive-self-improvement>) on "when AI builds itself," a trend that "might increase the risks of humans losing control over AI systems."

However, AI researchers were not impressed by Anthropic hamstringing its latest model's abilities.

"Anthropic's latest model will NOT help you if it thinks your ML research/ML engineering is interesting, and/or will secretly degrade its IQ so that the average engineer won't notice," AI research firm SemiAnalysis [tweeted](<https://x.com/SemiAnalysis_/status/2064482714149896431>).

"We are already seeing Anthropic's latest model's moderation filters our GPU inference research and programming," it added.

Other researchers accused Anthropic of using Fable 5 to "[shadowban](<https://x.com/rasbt/status/2064425877656543710>)," or quietly restrict the accounts, of AI researchers. According to the firm's system card, interventions limiting requests for "frontier LLM development" will "*not* be visible to the user."

This last concern, which could've effectively sabotaged anybody trying to train competing models by quietly bumping them down to less powerful models without their knowledge, proved controversial enough for Anthropic to change its mind.

"We’re changing Fable 5’s safeguards for frontier LLM development to make them visible," the company told *Wired* in a statement. "We made the wrong trade-off and we apologize for not getting the balance right."

"It felt like Anthropic was saying to the public, ‘We don't trust anybody else to do AI research," AI startup Prime Intellect research lead Will Brown told the publication. "We are the only ones who have to do AI research."

It all comes in the context Anthropic [calling for a global freeze on AI advances](<https://futurism.com/artificial-intelligence/anthropic-scared-calls-global-freeze-ai>) while discussing the dangers of "recursive self-improvement." In other words, the company is making a lot of noise about a sci-fi-sounding possibility: that AI will start to rapidly improve itself, potentially escaping the control of its human creators.

Beyond limiting its ability to develop AI tools, Fable 5's new safeguards also trigger when it encounters requests "related to cybersecurity, biology and chemistry, or distillation." Distillation is effectively using machine learning to train a "student" model on the behavior and reasoning of a "teacher" model, a practice that has sparked its fair share of controversy.

Anthropic has [already publicly griped](<https://futurism.com/artificial-intelligence/anthropic-deepseek-copying-ai>) about large-scale attempts to distill, or "extract" its underlying model — a [hypocritical stance](<https://futurism.com/future-society/google-copying-ai-permission>) given its indiscriminate scraping of rights-protected content on the web to train its AI in the first place.

**More on Anthropic:** [*Anthropic Scared, Calls for Global Freeze on AI Advances*](<https://futurism.com/artificial-intelligence/anthropic-scared-calls-global-freeze-ai>)

## Author
I've been at Futurism since 2017, where my role has evolved to encompass design, writing, and increasingly editing. I've always been fascinated by space exploration and advanced transportation, which I've leaned into by interviewing luminaries in those fields while closely following the dimensions of policy and regulation that allow next-generation projects to succeed -- or, sometimes, to fail. I'm also keenly interested in the effects of generative AI on society, policies, and democratic institutions, as well as clean energy, physics and biology, and the vagaries of tech leadership. My work for Futurism has been cited by publications including Ars Technica, Gizmodo, PC Magazine, Jalopnik, Fox News, and the New York Post. I spent my childhood living in locations including Manila, the Philippines, and Geneva, Switzerland, attended McGill University, and now live in Toronto, Canada. Before Futurism I worked at AskMen and a small photography studio. In my free time, I'm an avid gardener, foodie, and craft beer lover, as well as a maker of artisanal hot pepper sauces. I have a magnificent dog named Freida.

### Author social links  
[Bluesky](<https://bsky.app/profile/vtanger.bsky.social>)