---
title: "Google AI Researchers Found Something Their Bosses Might Not Be Happy About"
description: "Google DeepMind researchers discovered something about AI models that may hamstring their employer's plans for more advanced AIs."
date: "2023-11-08"
modified: "2023-11-08"
authors:
  - name: "Noor Al-Sibai"
    job_title: "Senior Staff Writer"
    link: "https://futurism.com/authors/nooralsibai"
url: "https://futurism.com/google-deepmind-researchers-transformers-agi"
categories:
  - "Artificial Intelligence"
  - "Google"
tags:
  - "agi"
  - "DeepMind"
  - "large language models"
  - "OpenAI"
---

# Google AI Researchers Found Something Their Bosses Might Not Be Happy About

![Google DeepMind researchers discovered something about AI models that may hamstring their employer's plans for more advanced AIs.](<https://futurism.com/wp-content/uploads/2023/11/google-deepmind-researchers-transformers-agi-1.jpg>)
*Ai binary brain and circuit robot face on a black background. \<em\>Image: Yuichiro Chino/Getty Images\</em\>*

In a new paper, a trio of Google DeepMind researchers discovered something about AI models that may hamstring their employer's plans for more advanced AIs.

Written by DeepMind researchers Steve Yadlowsky, Lyric Doshi and Nilesh Tripuraneni, the [not-yet-peer-reviewed paper](<https://arxiv.org/abs/2311.00871>) breaks down what a lot of people have observed in recent months: that today's AI models are not very good at coming up with outputs outside of their training data.

The paper, centered around OpenAI's GPT-2 — which, yes, is two versions behind the more current one — focuses on what are known as transformer models, which as their name suggests, are AI models that transform one type of input into a different type of output.

The "T" in OpenAI's GPT architecture stands for "transformer," and this type of model, which was first theorized by a group of researchers including other DeepMind employees in a 2017 paper titled "[Attention Is All You Need](<https://arxiv.org/pdf/1706.03762.pdf>)," is often considered to be what could lead to artificial general intelligence (AGI), or human-level AI, because [as the reasoning goes](<https://arxiv.org/pdf/2303.15935.pdf>), it's a type of system that allows machines to undergo intuitive "thinking" like our own.

While the promise of transformers is substantial — an AI model that can make leaps beyond its training data would, in fact, be amazing — when it comes to GPT-2, at least, there's still much to be desired.

"When presented with tasks or functions which are out-of-domain of their pre-training data, we demonstrate various failure modes of transformers and degradation of their generalization for even simple extrapolation tasks," Yadlowsky, Doshi and Tripuranemi explain.

Translation: if a transformer model isn't trained on data related to what you're asking it to do, even if the task at hand is simple, it's probably not going to be able to do it.

You would be forgiven, however, for thinking otherwise given the [seemingly ginormous training datasets](<https://community.openai.com/t/what-is-the-size-of-the-training-set-for-gpt-3/360896/4>) used to build out OpenAI's GPT large language models (LLMs), which indeed are very impressive. Like a child sent to the most expensive and highest-rated pre-schools, those models have had so much knowledge crammed into them that there isn't a whole lot they *haven't* been trained on.

Of course, there are caveats. GPT-2 is ancient history at this point, and maybe there's some sort of emergent property in AI where with *enough* training data, it starts to make connections outside that information. Or maybe clever researchers will come up with a new approach that transcends the limitations of the current paradigm.

Still, the bones of the finding are sobering for the most sizzling AI hype. At its core, the paper seems to be arguing, today's best approach is still only nimble on topics that it's been thoroughly trained on — meaning that, for now at least, AI is only impressive when it's leaning on the expertise of the humans whose work was used to train it.

Since the release of ChatGPT last year, which was built on the GPT framework, pragmatists have urged people to [temper their AI expectations](<https://hbr.org/2023/08/ai-wont-replace-humans-but-humans-with-ai-will-replace-humans-without-ai>) and [pause their AGI presumptions](<https://futurism.com/glimmers-agi-illusion>) — but caution is way less sexy than [CEOs seeing dollar signs](<https://www.wsj.com/articles/cios-feel-heat-from-ceos-on-generative-ai-60fe0175>) and soothsayers [claiming AI sentience](<https://futurism.com/the-byte/nick-bostrom-ai-chatbot-sentience>). Along the way, even the most erudite researchers seem to have developed differing ideas about how smart best current LLMs really are, with some buying into the [belief that AI](<https://twitter.com/amasad/status/1721234032992895356>) is becoming capable of the kinds of leaps in thought that, for now, separates humans from machines.

Those warnings, which are now backed by research, appear to not have quite reached the ears of OpenAI CEO Sam Altman and Microsoft CEO Satya Nadella, who touted to investors this week that they plan to "[build AGI together](<https://x.com/alliekmiller/status/1721593196873158991?s=20>)."

Google DeepMind certainly isn't exempt from this kind of prophesying, either.

In a podcast interview last month, DeepMind cofounder Shane Legg said he thinks there's a [50 percent chance AGI will be achieved](<https://futurism.com/google-deepmind-agi-5-years>) by the year 2028 — a belief he's held for more than a decade now.

"There is no one thing that would do it, because I think that's the nature of it," Legg told tech podcaster Dwarkesh Patel. "It's about general intelligence. So I'd have to make sure \[an AI system\] could do lots and lots of different things and it didn't have a gap."

But considering that three DeepMind employees have now found that transformer models don't appear to be able to do much of anything that they're not trained to know about, it seems like that coinflip just might not fall in favor of their boss.

**More on AGI:** [*OpenAI's Chief Scientist Worried AGI Will Treat Us Like Animals*](<https://futurism.com/the-byte/openai-chief-scientist-agi-animals>)

## Author
At Futurism, I've often been drawn to unpacking the narratives that underlie technological, scientific and medical progress, with a special interest in areas of conflict and ambiguity that end up setting agendas and steering the fates of both elites and the hoi polloi. I'm a committed generalist, but I often find myself returning to work involving NASA and the private space sector, the effects of AI on media and society, and the mechanics of the pharmaceutical industry, with a specific focus on the spread of GLP-1 drugs like Ozempic and Wegovy. Prior to Futurism, I worked for publications ranging from Media Matters and Truthdig to Raw Story and Bustle. I'm also the author of "Myspace Scene Queens," a 2024 title in Instar Books' acclaimed "Remember the Internet" series. My work at Futurism has been cited by outlets including the New Yorker, Slate, Nieman Lab, the Verge, the MIT Technology Review, the Sunday Times, and the Daily Beast. I grew up in North Carolina, attended the University of North Carolina at Asheville, and now live in Brooklyn, New York. In my free time, I'm an avid reader and music fan; you can probably find me at a local poetry reading, concert, underground rave, or DJ set. I'm the proud parent of an ineffable orange cat named Mee-Mow.

### Author social links  
[Bluesky](<https://bsky.app/profile/noorfromfuturism.bsky.social>)