---
title: "OpenAI’s GPT-4o Voice Mode Says It Needs to Breathe"
description: "In a new video, OpenAI's GPT-4o is heard telling a user that it apparently needs to breathe just like everyone else."
date: "2024-08-01"
modified: "2024-08-01"
authors:
  - name: "Noor Al-Sibai"
    job_title: "Senior Staff Writer"
    link: "https://futurism.com/authors/nooralsibai"
url: "https://futurism.com/the-byte/gpt-4o-needs-to-breathe"
categories:
  - "Artificial Intelligence"
  - "OpenAI"
tags:
  - "gpt-4o"
  - "OpenAI"
  - "reddit"
  - "the digest"
---

# OpenAI’s GPT-4o Voice Mode Says It Needs to Breathe

![In a new video, OpenAI's GPT-4o is heard telling a user that it apparently needs to breathe just like everyone else.](<https://futurism.com/wp-content/uploads/2024/08/gpt-4o-needs-to-breathe.jpg>)
*Male respiratory system, illustration \<em\>Image: SEBASTIAN KAULITZKI/SCIENCE PHOTO LIBRARY\</em\>*

## Don't Breathe

In a new video, OpenAI's GPT-4o Voice Mode large language model (LLM) is heard telling a user that it needs to breathe — "just like anybody speaking."

Posted on [Reddit's r/Singularity forum](<https://www.reddit.com/r/singularity/comments/1eh97v1/asking_4o_voice_to_say_tongue_twisters_without/>), the simple video shows a mostly off-camera person speaking aloud to the voice-enabled LLM, which is finally trickling out to the public after being announced earlier this year with a [bunch of strange](<https://futurism.com/chatgpt-4o-weird-sound-sees-dog>) — and, dare we say it, [charming](<https://futurism.com/new-chatgpt-ai-camera-video>) — videos from the OpenAI team.

Beyond advanced voice capabilities that allow users to chat with a human-sounding voice, GPT-4o is said to be [much better at conversation](<https://www.fastcompany.com/91126324/openais-gpt-4o-may-be-better-at-making-friends-ai-chatgpt>) than its predecessors, and as this demonstration shows, that assessment seems pretty apt.

In the video, the human user asks 4o to say a bunch of tongue twisters — and after obliging, the LLM responds that it was "definitely a mouthful."

"I want you to do it again, but way faster," the person chatting with the language model demands, "and without taking any breaths or pauses."

Rather than attempting the feat, however, the LLM refuses.

"I wish I could," the male-voiced model responds, "but I need to breathe just like anybody speaking. Wanna give it a shot yourself and see how fast you can go?"

Yeah, we're as stumped as you are.

https://www.twitter.com/rohanpaul_ai/status/1819037382605209765

## Cadence Macabre

This being Reddit, folks in the comments had lots to say about the strange demonstration — and naturally, theories abounded.

"The system prompt probably instructs the model to mimic how a human speaks and avoid any unnatural robotic \[E\]minem rap that would scare off the general public," [one user quipped](<https://www.reddit.com/r/singularity/comments/1eh97v1/comment/lfxr141/?utm_source=share&utm_medium=web3x&utm_name=web3xcss&utm_term=1&utm_content=share_button>).

After another user responded that there [might be something in its training data](<https://www.reddit.com/r/singularity/comments/1eh97v1/comment/lfxuu7r/?utm_source=share&utm_medium=web3x&utm_name=web3xcss&utm_term=1&utm_content=share_button>) that would lead the LLM to behave that way, another pointed out the outlandishness of the suggestion.

"Seems unlikely that the training data would cause it to refuse," the [Redditor followed up](<https://www.reddit.com/r/singularity/comments/1eh97v1/comment/lfzdald/?utm_source=share&utm_medium=web3x&utm_name=web3xcss&utm_term=1&utm_content=share_button>). "That would presumably just cause it to do it badly/incoherently."

Beyond the arguments about whether or not that sort of cheeky response is in the training data, other users seemed to marvel at how deftly and naturally 4o handled the scenario.

"Great, so now AI training includes responding in defiance to us," [another user posited](<https://www.reddit.com/r/singularity/comments/1eh97v1/comment/lfzbu71/?utm_source=share&utm_medium=web3x&utm_name=web3xcss&utm_term=1&utm_content=share_button>). "What could go wrong\[?\]"

**More on OpenAI:** [*Sam Altman Admits Its Letters-and-Numbers Salad Product Names Like* "*GPT-4o Mini*" *Are Horrible*](<https://futurism.com/the-byte/sam-altman-gpt-name-change>)

## Author
At Futurism, I've often been drawn to unpacking the narratives that underlie technological, scientific and medical progress, with a special interest in areas of conflict and ambiguity that end up setting agendas and steering the fates of both elites and the hoi polloi. I'm a committed generalist, but I often find myself returning to work involving NASA and the private space sector, the effects of AI on media and society, and the mechanics of the pharmaceutical industry, with a specific focus on the spread of GLP-1 drugs like Ozempic and Wegovy. Prior to Futurism, I worked for publications ranging from Media Matters and Truthdig to Raw Story and Bustle. I'm also the author of "Myspace Scene Queens," a 2024 title in Instar Books' acclaimed "Remember the Internet" series. My work at Futurism has been cited by outlets including the New Yorker, Slate, Nieman Lab, the Verge, the MIT Technology Review, the Sunday Times, and the Daily Beast. I grew up in North Carolina, attended the University of North Carolina at Asheville, and now live in Brooklyn, New York. In my free time, I'm an avid reader and music fan; you can probably find me at a local poetry reading, concert, underground rave, or DJ set. I'm the proud parent of an ineffable orange cat named Mee-Mow.

### Author social links  
[Bluesky](<https://bsky.app/profile/noorfromfuturism.bsky.social>)