---
title: "Amazon AGI Team Say Their AI Is Showing “Emergent Abilities”"
description: "A new Amazon AI model is, according to the researchers that built it, exhibiting incredible language abilities that it wasn't trained on. "
date: "2024-02-15"
modified: "2024-02-15"
authors:
  - name: "Noor Al-Sibai"
    job_title: "Senior Staff Writer"
    link: "https://futurism.com/authors/nooralsibai"
url: "https://futurism.com/the-byte/amazon-researchers-ai-emergent"
categories:
  - "Artificial Intelligence"
tags:
  - "agi"
  - "Amazon"
  - "llm"
  - "the digest"
---

# Amazon AGI Team Say Their AI Is Showing “Emergent Abilities”

![A new Amazon AI model is, according to the researchers that built it, exhibiting incredible language abilities that it wasn't trained on. ](<https://futurism.com/wp-content/uploads/2024/02/amazon-researchers-ai-emergent.jpg>)
*\<em\>Image: Getty / Futurism\</em\>*

## Training Day

A new Amazon AI model, according to the researchers who built it, is exhibiting language abilities that it wasn't trained on.

In a [not-yet-peer-reviewed academic paper](<https://arxiv.org/abs/2402.08093>), the team at Amazon AGI — which stands for "artificial general intelligence," or human-level AI — say their large language model (LLM) is exhibiting "state-of-the-art naturalness" at conversational text. Per the examples shared in the paper, the model does seem sophisticated.

As the paper indicates, the model was able to come up with all sorts of sentences that, according to criteria crafted with the help of an "expert linguist," showed it was making the types of language leaps that are natural in human language learners but have been difficult to obtain in AI.

Named "[Big Adaptive Streamable TTS with Emergent abilities](<https://amazon-ltts-paper.com/>)" or BASE TTS, the initial model was trained on 100,000 hours of "public domain speech data," 90 percent in English, to teach it how Americans talk. To test out how large models would need to be to show "emergent abilities," or abilities they were not trained on, the Amazon AGI team trained two smaller models, one on 1,000 hours of speech data and another on 10,000, to see which of the three — if any — exhibited the type of language naturalness they were looking for.

Interestingly enough, it was the 10,000-hour model — the Goldilocks of the three, if you will — that scored highest on the Amazon researchers' emergent abilities criteria list, which included things like the ability to understand punctuation, non-English words, and emotions.

## Speak Now

The middle model spat out sentences that would seem to human readers very natural, exhibiting the ability to transcribe non-words ("Shh, Lucy, shhh, we mustn’t wake your baby brother,” Tom whispered, as they tiptoed past the nursery") and even the kind of internetspeak many netizens use in text messages and spoken language alike ("She received an odd text from her brother: 'Emergency @ home; call ASAP! Mom & Dad are worried…#familymatters.'")

In the paper, whose international team of authors includes 18 AI experts, the Amazon AGI consortium pointed out that BASE TTS was never "explicitly" told to come up with its more surprising outputs.

"These sentences are designed to contain challenging tasks — parsing garden-path sentences, placing phrasal stress on long-winded compound nouns, producing emotional or whispered speech, or producing the correct phonemes for foreign words like “qi” or punctuations like “@” — none of which BASE TTS is explicitly trained to perform," the paper reads.

It's not AGI, of course — but these findings could regardless have implications on the path towards that goal, especially if it didn't need such a gigantic set of training data to get there.

**More on AI leaps:** [*AI Used to Resurrect Dead Dictator to Sway Election*](<https://futurism.com/the-byte/ai-resurrect-dead-dictator>)

## Author
At Futurism, I've often been drawn to unpacking the narratives that underlie technological, scientific and medical progress, with a special interest in areas of conflict and ambiguity that end up setting agendas and steering the fates of both elites and the hoi polloi. I'm a committed generalist, but I often find myself returning to work involving NASA and the private space sector, the effects of AI on media and society, and the mechanics of the pharmaceutical industry, with a specific focus on the spread of GLP-1 drugs like Ozempic and Wegovy. Prior to Futurism, I worked for publications ranging from Media Matters and Truthdig to Raw Story and Bustle. I'm also the author of "Myspace Scene Queens," a 2024 title in Instar Books' acclaimed "Remember the Internet" series. My work at Futurism has been cited by outlets including the New Yorker, Slate, Nieman Lab, the Verge, the MIT Technology Review, the Sunday Times, and the Daily Beast. I grew up in North Carolina, attended the University of North Carolina at Asheville, and now live in Brooklyn, New York. In my free time, I'm an avid reader and music fan; you can probably find me at a local poetry reading, concert, underground rave, or DJ set. I'm the proud parent of an ineffable orange cat named Mee-Mow.

### Author social links  
[Bluesky](<https://bsky.app/profile/noorfromfuturism.bsky.social>)