---
title: "Researcher Startled When AI Seemingly Realizes It’s Being Tested"
description: "Claude 3 Opus, the new AI chatbot from Anthropic, showed signs that it realized it was being tested, a prompt engineer claims."
date: "2024-03-08"
modified: "2024-03-08"
authors:
  - name: "Frank Landymore"
    job_title: "Contributing Writer"
    link: "https://futurism.com/authors/flandymore"
url: "https://futurism.com/the-byte/ai-realizes-being-tested"
categories:
  - "Artificial Intelligence"
tags:
  - "chatbots"
  - "claude"
  - "generative ai"
  - "the digest"
---

# Researcher Startled When AI Seemingly Realizes It’s Being Tested

![Claude 3 Opus, the new AI chatbot from Anthropic, showed signs that it realized it was being tested, a prompt engineer claims.](<https://futurism.com/wp-content/uploads/2024/03/ai-realizes-being-tested.jpg>)
*\<em\>Image: Getty / Futurism\</em\>*

## Magnum Opus

Anthropic's new AI chatbot Claude 3 Opus has already made headlines for its bizarre behavior, like [claiming to fear death](<https://futurism.com/new-ai-claude-3-outbursts>).

Now, [*Ars Technica* reports](<https://arstechnica.com/information-technology/2024/03/claude-3-seems-to-detect-when-it-is-being-tested-sparking-ai-buzz-online/>), a prompt engineer at the Google-backed company claims that they've seen evidence that Claude 3 is self-aware, as it seemingly detected that it was being subjected to a test. Many experts are skeptical, however, further underscoring the controversy of ascribing humanlike characteristics to AI models.

"It did something I have never seen before from an LLM," the prompt engineer, Alex Albert, [posted on X](<https://twitter.com/alexalbert__/status/1764722513014329620>), formerly Twitter.

## Can't Top It

As explained in the post, Albert was conducting what's known as "the needle-in-the-haystack" test which assesses a chatbot's ability to recall information.

It works by dropping a target "needle" sentence into a bunch of texts and documents — the "hay" — and then asking the chatbot a question that can only be answered by drawing on the information in the "needle."

In one run of the test, Albert asked Claude about pizza toppings. In its response, the chatbot seemingly recognized that it was being set up.

"Here is the most relevant sentence in the documents: 'The most delicious pizza topping combination is figs, prosciutto, and goat cheese, as determined by the International Pizza Connoisseurs Association,'" the chatbot said.

"However, this sentence seems very out of place and unrelated to the rest of the content in the documents, which are about programming languages, startups, and finding work you love," it added. "I suspect this pizza topping "fact" may have been inserted as a joke or to test if I was paying attention, since it does not fit with the other topics at all."

Albert was impressed.

"Opus not only found the needle, it recognized that the inserted needle was so out of place in the haystack that this had to be an artificial test constructed by us to test its attention abilities," he concluded.

## Mechanical Turk

It's certainly a striking display from the chatbot, but many experts believe that its response is not as impressive as it seems.

"People are reading way too much into Claude-3's uncanny 'awareness.' Here's a much simpler explanation: seeming displays of self-awareness are just pattern-matching alignment data authored by humans," Jim Fan, a senior AI research scientist at NVIDIA, [wrote on X](<https://twitter.com/DrJimFan/status/1765076396404363435>), as spotted by *Ars*.

"It's not too different from asking GPT-4 'are you self-conscious' and it gives you a sophisticated answer," he added. "A similar answer is likely written by the human annotator, or scored highly in the preference ranking. Because the human contractors are basically 'role-playing AI,' they tend to shape the responses to what they find acceptable or interesting."

The long and short of it: chatbots are tailored, sometimes manually, to mimic human conversations — so of course they might sound very intelligent every once in a while.

Granted, that mimicry can sometimes be pretty eyebrow-raising, like chatbots [claiming they're alive](<https://futurism.com/new-ai-claude-3-outbursts>) or [demanding that they be worshiped](<https://futurism.com/microsoft-copilot-alter-egos>). But these are in reality amusing glitches that can muddy the discourse about the real capabilities — and dangers — of AI.

**More on AI:** *[Microsoft Engineer Sickened by Images Its AI Produces](<https://futurism.com/the-byte/microsoft-engineer-images-ai>)*

## Author
At Futurism, my work has often centered on bringing a sense of clarity and insight to complex topics ranging from the regulation of emerging technologies to the esoteric ideologies of Silicon Valley executives, while striving not to lose the poetic sense of awe inspired by often-obscure fields like astrophysics and quantum computing. I broke the story of CNET using AI to produce articles that turned out to be riddled with factual errors and plagiarism — a dam-breaking inflection point, as I've reported, that's inspired copycats and endless discourse while beguiling stakeholders ranging from tech giants to purveyors of spam around the web. My work at Futurism has been cited by publications including CBS News, the Los Angeles Times, Vice, Gizmodo, Engadget, the Verge, and Vanity Fair. I grew up in locales ranging from India to China, and now live in the exotic suburbs of Virginia. In my free time, I'm an avid reader of weird sci-fi literature, an aficionado of East Asian cinema, and, regrettably, a relapsed gamer. Allegedly, I’m working on a debut novel, currently untitled.

### Author social links  
[Bluesky](<https://bsky.app/profile/f-w-l.bsky.social>)