---
title: "OpenAI Safety Worker Quit Due to Losing Confidence Company “Would Behave Responsibly Around the Time of AGI”"
description: "An OpenAI safety worker quit his job, arguing in an online forum that he had lost confidence that the Sam Altman-led company will \"behave responsibly around the time of [artificial general intelligence],\" the theoretical point at which an AI can outperform a human."
date: "2024-05-13"
modified: "2024-05-13"
authors:
  - name: "Victor Tangermann"
    job_title: "Senior Editor"
    link: "https://futurism.com/authors/victor"
url: "https://futurism.com/openai-safety-worker-quit-confidence-agi"
categories:
  - "Artificial Intelligence"
  - "OpenAI"
tags:
  - "agi"
  - "artificial general intelligence"
  - "OpenAI"
  - "sam altman"
---

# OpenAI Safety Worker Quit Due to Losing Confidence Company “Would Behave Responsibly Around the Time of AGI”

![An OpenAI safety worker quit his job, arguing in an online forum that he had lost confidence that the Sam Altman-led company will "behave responsibly around the time of \[artificial general intelligence\]," the theoretical point at which an AI can outperform a human.](<https://futurism.com/wp-content/uploads/2024/05/openai-safety-worker-quits-confidence-agi.jpg>)
*\<em\>Image: Getty / Futurism\</em\>*

An OpenAI safety worker quit his job, [arguing in an online forum](<https://www.lesswrong.com/users/daniel-kokotajlo>) that he had lost confidence that the Sam Altman-led company will "behave responsibly around the time of \[artificial general intelligence\]," the theoretical point at which an AI can outperform a human.

As [*Business Insider* reports](<https://www.businessinsider.com/openai-safety-researchers-quit-superalignment-sam-altman-chatgpt-2024-5>), researcher Daniel Kokotajlo, a philosophy PhD student who worked in OpenAI's governance team, left the company last month.

In several followup posts on the forum LessWrong, Kokotajlo explained his "disillusionment" that led to him quitting, which was related to a growing call to put a pause on research that could eventually lead to the establishment of AGI.

It's a heated debate, with experts long [warning of](<https://futurism.com/group-temporary-pause-ai-more-advanced-than-gpt-4>) the [potential dangers](<https://futurism.com/ai-expert-bomb-datacenters>) of an AI that exceeds the cognitive capabilities of humans. Last year, over 1,100 artificial intelligence experts, CEOs, and researchers — including SpaceX CEO Elon Musk — signed an [open letter](<https://futureoflife.org/open-letter/pause-giant-ai-experiments/>) calling for a six-month moratorium on "AI experiments."

"I think most people pushing for a pause are trying to push against a 'selective pause' and for an actual pause that would apply to the big labs who are at the forefront of progress," Kokotajlo [wrote](<https://www.lesswrong.com/posts/3LuZm3Lhxt6aSpMjF/ai-regulation-is-unsafe?commentId=nDgJuyyhrs3KpyekS>).

However, he argued that such a "selective pause" would end up not applying to the "big corporations that most need to pause."

"My disillusionment about this is part of why I left OpenAI," he concluded.

Kokotajlo quit roughly two months after research engineer William Saunders left the company as well.

The Superalignment team, which Saunders was part of at OpenAI for three years, was cofounded by computer scientist and former OpenAI chief scientist Ilya Sutskever and his colleague Jan Leike. It's tasked with ensuring that "AI systems much smarter than humans follow human intent," according to OpenAI's website.

"Superintelligence will be the most impactful technology humanity has ever invented, and could help us solve many of the world’s most important problems," the company's [description](<https://openai.com/superalignment/>) of the team reads. "But the vast power of superintelligence could also be very dangerous, and could lead to the disempowerment of humanity or even human extinction."

Instead of having a "solution for steering or controlling a potentially superintelligent AI, and preventing it from going rogue," the company is hoping that "scientific and technical breakthroughs" could lead to an equally superhuman alignment tool that can keep systems that are "much smarter than us" in check.

But given Saunders' departure, it seems like not everybody on the Superalignment team were themselves aligned on the company's ability to police an eventual AGI.

The debate surrounding the dangers of an unchecked superintelligent AI may have played a role in the firing and eventual rehiring of CEO Sam Altman last year. Sutskever, who used to sit on the original board of OpenAI's non-profit entity, [reportedly disagreed](<https://www.theinformation.com/articles/before-openai-ousted-altman-employees-disagreed-over-ai-safety>) with Altman on the topic of AI safety before Altman was ousted, and was later kicked off the board.

To be clear, all of this is still an entirely theoretical discussion. Despite [plenty](<https://futurism.com/artificial-superintelligence-agi-2027-goertzel>) [of](<https://futurism.com/google-deepmind-agi-5-years>) [predictions](<https://futurism.com/elon-musk-ai-prediction>) by experts that AGI is only a matter of years away, there's no guarantee that we'll ever reach a point at which an AI could outperform humans.

But if they do, it'll raise an incredibly important question: how do we ensure AGI systems don't go rogue if they're inherently more capable than us?

And not everybody is confident in OpenAI's ability and longterm commitment to controlling AGI, with the company risking becoming too big to be effectively regulated, as Kokotajlo argued.

**More on OpenAI:** *[OpenAI Mocked for Issuing Infringement Claim Over Its Logo While Scraping the Entire Web to Train AI Models](<https://futurism.com/the-byte/openai-logo-infringement-claim>)*

## Author
I've been at Futurism since 2017, where my role has evolved to encompass design, writing, and increasingly editing. I've always been fascinated by space exploration and advanced transportation, which I've leaned into by interviewing luminaries in those fields while closely following the dimensions of policy and regulation that allow next-generation projects to succeed -- or, sometimes, to fail. I'm also keenly interested in the effects of generative AI on society, policies, and democratic institutions, as well as clean energy, physics and biology, and the vagaries of tech leadership. My work for Futurism has been cited by publications including Ars Technica, Gizmodo, PC Magazine, Jalopnik, Fox News, and the New York Post. I spent my childhood living in locations including Manila, the Philippines, and Geneva, Switzerland, attended McGill University, and now live in Toronto, Canada. Before Futurism I worked at AskMen and a small photography studio. In my free time, I'm an avid gardener, foodie, and craft beer lover, as well as a maker of artisanal hot pepper sauces. I have a magnificent dog named Freida.

### Author social links  
[Bluesky](<https://bsky.app/profile/vtanger.bsky.social>)