---
title: "Astounding AI Guesses What You Look Like Based on Your Voice"
description: "By analyzing only a short audio clip of a person's voice, this artificial intelligence reconstructs what they might look like in real life."
date: "2019-05-28"
modified: "2019-05-31"
authors:
  - name: "Jon Christian"
    job_title: "Executive Editor"
    link: "https://futurism.com/authors/jonc"
url: "https://futurism.com/the-byte/ai-guesses-appearance-voice"
categories:
  - "Artificial Intelligence"
tags:
  - "artificial intelligence"
  - "facial recognition"
  - "the digest"
  - "voice recognition"
---

# Astounding AI Guesses What You Look Like Based on Your Voice

![By analyzing only a short audio clip of a person's voice, this artificial intelligence reconstructs what they might look like in real life.](<https://futurism.com/wp-content/uploads/2019/05/terrifying-ai-guesses-what-you-look-like-based-on-your-voice.png>)
*\<em\>Image: Wen et al./Victor Tangermann\</em\>*

## Vox Humana

A new artificial intelligence created by researchers at the Massachusetts Institute of Technology pulls off a staggering feat: by analyzing only a short audio clip of a person's voice, it reconstructs what they might look like in real life.

The AI's results aren't perfect, but they're pretty good — a remarkable and somewhat terrifying example of how a sophisticated AI can make incredible inferences from tiny snippets of data.

## Biometric Characteristics

In a paper [published this week](<https://arxiv.org/abs/1905.09773>) to the preprint server *arXiv,* the team describes how it used a deep network architecture, trained by videos from YouTube and elsewhere online, to analyze short voice clips and reconstruct what the speaker might look like.

In practice, the Speech2Face algorithm seems to have an uncanny knack for spitting out rough likenesses of people based on nothing but their speaking voices.

![](<https://futurism.com/wp-content/uploads/2019/05/face1.png>)

## Face/Off

The MIT research isn't the first to recreate a speaker's physical characteristics based on voice recordings. Researchers at Carnegie Mellon University recently published [a paper](<https://arxiv.org/pdf/1905.10604.pdf>) on a similar algorithm, which they [presented at the World Economic Forum](<https://www.weforum.org/agenda/2018/01/catch-criminal-milliseconds-audio-rita-singh-carnegie/>) last year.

The MIT team urges caution on the project's [GitHub page](<https://speech2face.github.io/>), acknowledging that the tech raises worrisome questions about privacy and discrimination.

"Although this is a purely academic investigation, we feel that it is important to explicitly discuss in the paper a set of ethical considerations due to the potential sensitivity of facial information," they wrote, suggesting that "any further investigation or practical use of this technology will be carefully tested to ensure that the training data is representative of the intended user population."

*Editor's note: This story mistakenly identified Speech2Voice as a Carnegie Mellon University project, not an MIT one. It has also been updated with technical details about the MIT project and background about previous work at Carnegie Mellon.*

**READ MORE:** [Speech2Face: Learning the Face Behind a Voice](<https://arxiv.org/abs/1905.09773>) \[*arXiv*\]

**More on neural networks:** *[A Neural Net Hooked Up to a Monkey Brain Spat Out Bizarre Images](<https://futurism.com/neural-net-monkey-brain-bizarre-images>)*

## Author
I'm responsible for editing, assigning, and scouting at Futurism, as well as occasionally writing for the site. That work involves keeping an eye on a wide range of narratives and issues, but over the past few years I've become increasingly interested in how AI is shaping the future of media, the web, and information itself. My day-to-day often involves collaborating on reporting and commentary projects bylined by my colleagues, but I try to find time to do my own reporting as well; stories I've broken for Futurism have been cited by publications including the New York Times, the Washington Post, the New Yorker, Wired, Ars Technica, New York Magazine, the Columbia Journalism Review, Bloomberg, and more. I grew up in Southern Vermont and attended Vanderbilt University. Prior to Futurism, I contributed to outlets including the Boston Globe, Vice, Wired, Slate, the Atlantic, the Outline, Ars Technica and more, and did stints in farming and food service. I currently live in Brooklyn, New York, and in my free time I enjoy pinball, word games, cycling, running, and music production.

### Author social links  
[Bluesky](<https://bsky.app/profile/jonchristian.net>)