---
title: "OpenAI’s GPT-4 Just Smoked Basically Every Test and Exam Anyone’s Ever Taken"
description: "OpenAI has released a bunch of stats about its even-more-powerful new language model called GPT-4 — and we're slightly freaked."
date: "2023-03-14"
modified: "2023-03-14"
authors:
  - name: "Noor Al-Sibai"
    job_title: "Senior Staff Writer"
    link: "https://futurism.com/authors/nooralsibai"
url: "https://futurism.com/the-byte/gpt-4-exam-scores"
categories:
  - "Artificial Intelligence"
  - "OpenAI"
tags:
  - "gpt-4"
  - "large language models"
  - "OpenAI"
  - "the digest"
---

# OpenAI’s GPT-4 Just Smoked Basically Every Test and Exam Anyone’s Ever Taken

![Hot on the heels of the GPT-4 drop, OpenAI has released a bunch of stats about its even-more-powerful new language model — and we're slightly freaked.](<https://futurism.com/wp-content/uploads/2023/03/gpt-4-exam-scores.jpg>)
*\<em\>Image: Jakub Porzycki/NurPhoto via Getty Images\</em\>*

## Freak Out

OpenAI's GPT-4 is [officially here](<https://openai.com/research/gpt-4>) — and the numbers speak for themselves.

Hot on the [heels of its announcement](<https://twitter.com/sama/status/1635687853324902401>), OpenAI has released a bunch of stats about its even-more-powerful new large language model — and reader, we're both spooked and skeptical in equal measures.

According to a [new white paper](<https://cdn.openai.com/papers/gpt-4.pdf>), the algorithm got incredibly good scores on a number of exams including the Bar, the LSATs, the SAT's Reading and Math tests, and the GRE.

To put these high scores in perspective, it's important to look at the average scores for all the exams GPT-4 appears to have aced. For instance, the LLM got a 163 out of 180 on the LSAT, which is more than ten points higher than the median score of 152 ([per the *Princeton Review*](<https://www.princetonreview.com/law-school-advice/lsat-scores#:~:text=Your%20LSAT%20score%20is%20the,law%20schools%20you%20are%20considering.>)) and, perhaps even more remarkably, almost twice as good as its predecessor, GPT-3.

![](<https://futurism.com/wp-content/uploads/2023/03/Screen-Shot-2023-03-14-at-3.35.57-PM.png>)

## Limiting Factors

While these stats — which, to be very clear, were released by OpenAI itself and were undoubtedly tailored to make the LLM look as impressive as possible — are indeed stunning, the firm also admitted that its latest LLM is still suffering from the same drawbacks as its predecessors.

"Despite its capabilities, GPT-4 has similar limitations as earlier GPT models," [OpenAI noted on its website](<https://openai.com/research/gpt-4>). "Most importantly, it still is not fully reliable (it 'hallucinates' facts and makes reasoning errors)."

"Great care should be taken when using language model outputs," the AI firm added, "particularly in high-stakes contexts, with the exact protocol (such as human review, grounding with additional context, or avoiding high-stakes uses altogether) matching the needs of a specific use-case."

## Mixed Reviews

What is clear, if nothing else, is that OpenAI is racing ahead with the release of its LLMs — GPT-3 was [released in the summer of 2020](<https://www.cnbc.com/2020/07/23/openai-gpt3-explainer.html>); GPT 3.5, the update that gave the world ChatGPT, [dropped on the first of December of last year](<https://techcrunch.com/2022/12/01/while-anticipation-builds-for-gpt-4-openai-quietly-releases-gpt-3-5/>), and now, just three-ish months later, GPT-4.

While we're still waiting to find out about GPT-4's full capabilities, it's pretty obvious at this point that there's a lot of growing momentum — and financial interest — in the AI space.

If you want to give the new model a spin, it's a lot easier than you might think: [Microsoft has already confirmed](<https://blogs.bing.com/search/march_2023/Confirmed-the-new-Bing-runs-on-OpenAI%E2%80%99s-GPT-4>) that it's been using GPT-4 all along for its Bing AI search assistant.

**More on OpenAI:** [*OpenAI Confused by Why People Are So Impressed With ChatGPT*](<https://futurism.com/the-byte/openai-confused-people-impressed-chatgpt>)

## Author
At Futurism, I've often been drawn to unpacking the narratives that underlie technological, scientific and medical progress, with a special interest in areas of conflict and ambiguity that end up setting agendas and steering the fates of both elites and the hoi polloi. I'm a committed generalist, but I often find myself returning to work involving NASA and the private space sector, the effects of AI on media and society, and the mechanics of the pharmaceutical industry, with a specific focus on the spread of GLP-1 drugs like Ozempic and Wegovy. Prior to Futurism, I worked for publications ranging from Media Matters and Truthdig to Raw Story and Bustle. I'm also the author of "Myspace Scene Queens," a 2024 title in Instar Books' acclaimed "Remember the Internet" series. My work at Futurism has been cited by outlets including the New Yorker, Slate, Nieman Lab, the Verge, the MIT Technology Review, the Sunday Times, and the Daily Beast. I grew up in North Carolina, attended the University of North Carolina at Asheville, and now live in Brooklyn, New York. In my free time, I'm an avid reader and music fan; you can probably find me at a local poetry reading, concert, underground rave, or DJ set. I'm the proud parent of an ineffable orange cat named Mee-Mow.

### Author social links  
[Bluesky](<https://bsky.app/profile/noorfromfuturism.bsky.social>)