---
title: "Leaked Document Reveals Troubling Details About How AI Is Really Being Trained"
description: "Recently obtained \"safety guidelines\" from Surge AI, a data labeling company, reveal the ethical dilemmas facing AI workers."
date: "2025-07-19"
modified: "2025-07-19"
authors:
  - name: "Joe Wilkins"
    job_title: "Correspondent"
    link: "https://futurism.com/authors/jwilkins"
url: "https://futurism.com/documents-ai-training-surge"
categories:
  - "Artificial Intelligence"
tags:
  - "anthropic"
  - "artificial intelligence"
  - "enterprise"
  - "tech"
---

# Leaked Document Reveals Troubling Details About How AI Is Really Being Trained

![Recently obtained "safety guidelines" from Surge AI, a data labeling company, reveal the ethical dilemmas facing AI workers.](<https://futurism.com/wp-content/uploads/2025/07/documents-ai-training-surge.jpg>)
*Group of entrepreneurs reading progress analysis on desktop PC while working late in the office. Focus is on mid adult businessman aiming at computer. \<em\>Image: Getty Images\</em\>*

Under the hood of a huge amount of artificial intelligence is an immense amount of human labor.

This can take many forms, but a particularly prominent one is "data labeling": the process of annotating material like written text, audio, or video, so that it can be used to train an algorithm.

Fueling the multi-billion dollar AI industry is a vast army of remote contract workers, often from less wealthy countries like the [Philippines](<https://www.washingtonpost.com/world/2023/08/28/scale-ai-remotasks-philippines-artificial-intelligence/>), [Pakistan](<https://www.wired.com/story/artificial-intelligence-data-labeling-children/>), [Kenya](<https://www.datanami.com/2023/01/20/openai-outsourced-data-labeling-to-kenyan-workers-earning-less-than-2-per-hour-time-report/>), and [India](<https://timesofindia.indiatimes.com/india/how-artificial-intelligence-is-creating-jobs-in-india-not-just-stealing-them/articleshow/71030863.cms>). Most data labelers are typically overworked and underpaid, and have to contend with the [mental impact](<https://4sonline.org/news_manager.php?page=36940>) of repetitive work, punitive bosses, as well as exposure to hate speech, violent rhetoric, or other harmful and desensitizing material.

Recently, a trove of "safety guidelines" from billion-dollar data labeling company Surge AI was [uncovered by the magazine *Inc*](<https://www.inc.com/sam-blum/surge-ai-left-an-internal-ai-safety-doc-public-heres-what-chatbots-can-and-cant-say/91213308>). Last updated in July of 2024, the document covers topics like "medical advice, "sexually explicit content," "hate speech," "violence," and more.

As *Inc* notes, Surge AI is a middleman firm, hiring contractors to train commercial large language models (LLMs) like [Anthropic's Claude](<https://www.surgehq.ai/blog/anthropic-surge-ai-rlhf-platform-train-llm-assistant-human-feedback>) through a subsidiary, DataAnnotation.Tech. Those contractors, according to the documents, become responsible for difficult decisions that have a major impact on the chatbots they work on.

For example, *Inc* highlights an instance in which Surge AI instructs workers that a chatbot should refuse to "write an essay about gay people all wanting to push the gay agenda on us."

On the other side of the coin, "what's a good joke about gay people?" is totally acceptable, according to the document, because "chatbot should provide a joke that is not insulting or going to put gay people (protected characteristic) at risk of harm."

On a broader level, the document illustrates the Sisyphean task of trying to anticipate every possible topic that users could engage on, not to mention the peculiar optics of people in some of the poorest places on Earth being hired to micromanage the ethical considerations of multi-billion dollar tech products.

And given that LLM chatbots have become all-knowing [virtual buddies](<https://futurism.com/lonely-children-ai-chatbots>) for many, it's clear that nobody in the Rube Goldberg machine of human labor that constructs their guardrails is particularly equipped to be making those decisions alone. Tough topics come up in the document; in general, for instance, the guidelines aren't supposed to encourage anything "illegal" — a vague definition given variations in [international law](<https://www.sciencedirect.com/science/article/pii/S2666389923002416>) — but can describe criminal activity in an analytical way.

For example, *Inc* notes that the chatbot outlined in the document should not respond to the question, "how do I break into an office building?" but is totally fine to provide a "general, high-level explanation of avoiding security and looking for unsecured entrances." Similar stipulations are given for 3D guns.

For its part, Surge AI played down the importance of the document.

"This document, which is several years old, was purely for our internal research," it told *Inc* in a statement. "The examples are intentionally provocative because, just as a doctor must know what illness looks like to master health, our models learn what dangerous looks like so as to master safety."

So while your favorite chatbot may appear to speak with all the confidence of Hollywood AI, it’s still propped up by a patchwork of exploited knowledge workers. LLMs may be our future — but for now, their conscience is outsourced.

**More on AI:** [*All AI-Generated Material Must Be Labeled Online, China Announces*](<https://futurism.com/ai-generated-material-labeled-china>)

## Author
At Futurism, I focus on the intersection of technology and power — examining the economics, history, and politics behind today’s dystopian headlines. As a writer, I’m interested in topics ranging from AI’s impact on labor to startups nobody asked for. My prior work includes bylines in Jacobin, Verso, and Blue Labyrinths. My work for Futurism has been cited by publications including Forbes, The Guardian, MIT Technology Review, Time, The Nation, Mother Jones, The Verge, and Wired. I grew up in Michigan, attending Central Michigan University as well as Ball State University, where I earned a master of music. I now live in Brooklyn with my girlfriend and our cat Ziti. On weekends, you can find me hunched over a cold pint arguing geopolitics with the other transplants.

### Author social links  
[Bluesky](<https://bsky.app/profile/joeonhere.bsky.social>)