---
title: "OpenAI’s Video-Generating AI Is “Doomed to Failure,” Says Meta’s Top AI Scientist"
description: "Yann LeCun pulled no punches when he described OpenAI's approach to creating a \"world simulator\" through its Sora AI as a bad idea."
date: "2024-02-22"
modified: "2024-02-22"
authors:
  - name: "Frank Landymore"
    job_title: "Contributing Writer"
    link: "https://futurism.com/authors/flandymore"
url: "https://futurism.com/the-byte/openai-video-ai-doomed-meta-scientist"
categories:
  - "Artificial Intelligence"
  - "OpenAI"
tags:
  - "generative ai"
  - "sora"
  - "the digest"
  - "yann lecun"
---

# OpenAI’s Video-Generating AI Is “Doomed to Failure,” Says Meta’s Top AI Scientist

![Yann LeCun pulled no punches when he described OpenAI's approach to creating a "world simulator" through its Sora AI as a bad idea.](<https://futurism.com/wp-content/uploads/2024/02/openai-video-ai-doomed-meta-scientist.jpg>)
*\<em\>Image: Kent Nishimura via Getty / Futurism\</em\>*

## Pixel Imperfect

Sora, OpenAI's new AI model for generating video, has become the talk of the town since [its release last week](<https://futurism.com/openai-sora-video-generator>). But Yann LeCun, chief AI scientist at Meta, doesn't think the much-hyped text-to-video model is all that.

In particular, LeCun takes issue with OpenAI's [claims](<https://openai.com/research/video-generation-models-as-world-simulators>) that its work with Sora will eventually [enable the building](<https://futurism.com/openai-sora-ai-simulate-worlds>) of "general purpose simulators of the physical world." If that's the case, LeCun argues, its approach to creating a "world simulator" is dead wrong.

"Modeling the world for action by generating pixels is as wasteful and doomed to failure as the largely-abandoned idea of 'analysis by synthesis,'" he wrote in a [post on X](<https://twitter.com/ylecun/status/1759486703696318935>), formerly Twitter.

## Generation Complication

LeCun is one of the so-called godfathers of AI — and perhaps the bluntest and most outspoken one at that. While the other two [godfathers lament](<https://futurism.com/the-byte/godfather-ai-take-over>) [what they've unleashed](<https://futurism.com/the-byte/ai-inventor-regret>), LeCun has [pushed on](<https://futurism.com/the-byte/godfather-ai-stop-freaking-out>) with his work at Meta, never afraid to criticize his competitors.

With his comments here, he's referring to an age-old debate in machine learning between generative models and discriminative models. LeCun believes that the former approach, generating pixels "from explanatory latent variables," is too inefficient, and can't adequately deal with the uncertainty that arises from these complex predictions in a 3D space.

In layman's terms, he's arguing that these models are trying to "infer" too many details that aren't relevant — sort of like trying to calculate the trajectory of a soccer ball by trying understand how every material it's made of comes into play instead of just focusing on stuff like its mass and velocity.

"There is nothing wrong with that if your purpose is to actually generate videos," he said in a [reply](<https://twitter.com/ylecun/status/1759495639224688774>) to his post. "But if your purpose is to understand how the world works, it's a losing proposition."

## The Alternative

LeCun concedes that, by and large, the generative approach has worked with large language models like ChatGPT so far "because text is discrete with a finite number of symbols." But if you're going to simulate the world like Sora is supposed to, you're dealing with much more than just a couple of characters.

Competing with OpenAI's approach, LeCun has been working on his own model at Meta called the Video Joint Embedding Predictive Architecture (V-JEPA), which was unveiled last week.

"Unlike generative approaches that try to fill in every missing pixel," Meta claims in a [blog post](<https://ai.meta.com/blog/v-jepa-yann-lecun-ai-model-video-joint-embedding-predictive-architecture/>), "V-JEPA has the flexibility to discard unpredictable information, which leads to improved training and sample efficiency by a factor between 1.5x and 6x."

LeCun's work may not get all the hype that OpenAI's products do with their flashy image and text generation, but it is interesting to see such a prominent AI researcher diverging from the same old approaches currently being developed by OpenAI and its slew of imitators.

**More on AI:** *[ChatGPT Appears to Have Lost Its Mind Last Night](<https://futurism.com/chatgpt-lost-mind>)*

## Author
At Futurism, my work has often centered on bringing a sense of clarity and insight to complex topics ranging from the regulation of emerging technologies to the esoteric ideologies of Silicon Valley executives, while striving not to lose the poetic sense of awe inspired by often-obscure fields like astrophysics and quantum computing. I broke the story of CNET using AI to produce articles that turned out to be riddled with factual errors and plagiarism — a dam-breaking inflection point, as I've reported, that's inspired copycats and endless discourse while beguiling stakeholders ranging from tech giants to purveyors of spam around the web. My work at Futurism has been cited by publications including CBS News, the Los Angeles Times, Vice, Gizmodo, Engadget, the Verge, and Vanity Fair. I grew up in locales ranging from India to China, and now live in the exotic suburbs of Virginia. In my free time, I'm an avid reader of weird sci-fi literature, an aficionado of East Asian cinema, and, regrettably, a relapsed gamer. Allegedly, I’m working on a debut novel, currently untitled.

### Author social links  
[Bluesky](<https://bsky.app/profile/f-w-l.bsky.social>)