> Llama 3 is an open-weights AI model from Meta, released 18 Apr 2024. Index score 98 (GPT-4 = 100), rank 188 of 229 ranked models.

[Models](https://themodelindex.org/models/) /[Meta](https://themodelindex.org/labs/meta/)

# Llama 3

Meta

Released 18 Apr 2024

Open

Date verified

[Compare](https://themodelindex.org/compare/llama-3-vs-gpt-4o/) [Announcement](https://ai.meta.com/blog/meta-llama-3/)

9886–110

Index score

#188 of 229 models

1/ 8

Core tests taken

1 of 6 areas · plus 5 anchor tests

−9

Versus the frontier at release

best then: GPT-4 Turbo (Apr 2024)

Replaced by [Llama 3.1-70B](https://themodelindex.org/models/llama-3-1-70b/)

–

Price per million tokens

no public API price found

–

Context window

tokens it reads at once

## Where it sits · on the frontier

_Chart: Model Index score of every scored model since late 2022, with the best closed and best open model over time_

GPT-4 reached this score 13 months earlier.

## Test results · by area · tick = best by any model

Show 4 anchor tests

Reasoning & knowledge

evidence

GPQA Diamond

41%

Humanity's Last Exam

no result

Maths

evidence

FrontierMath T1-3

no result

Coding

no evidence yet

SWE-bench Verified

no result

Terminal-Bench

no result

Agents

evidence

APEX-Agents

no result

Novel problems

no evidence yet

ARC-AGI-2

no result

Long tasks

no evidence yet

METR time horizon

no result

Independent results from the Epoch AI benchmarking hub, best reasoning setting. The core tests are shown here; anchor tests also inform the score and can be shown above.

## What others say · rank on each leaderboard

The Model Index

independent tests, our method

#188

of 219 · 98

Epoch Capabilities Index

overall capability from many benchmarks

#157

of 239 · 122.9

LMArena · Creative writing

creative writing

#267

of 386 · 1254

LMArena · Maths

maths questions in chat

#265

of 373 · 1258

LMArena · Text

overall chat quality, judged by people

#278

of 388 · 1276

LMArena · Coding

coding questions in chat

#279

of 383 · 1306

Further left is better. Leaderboards measure different things: LMArena is people voting blind between two answers, the others are test scores. Click any row for the source.

## Best at · top three among current models

Replaced by Llama 3.1-70B; it no longer leads any category among current models.

Where this model places in the top three of today's models, on category leaderboards and on our core tests.

## Facts and sources

Released

18 Apr 2024

Announcement

[ai.meta.com ↗](https://ai.meta.com/blog/meta-llama-3/)

Release entry

Llama 3

Scored as

Llama 3-70B

Artificial Analysis index

–

Output speed

–

Hugging Face

not linked

## Meta releases · around this one

Llama 2

Open

18 JUL 2023

index 88

Code Llama

Open

24 AUG 2023

Llama 3

Open

18 APR 2024

index 98

Llama 3.1

Open

23 JUL 2024

index 108

Llama 3.2

Open

25 SEP 2024

index 103

---
Source: https://themodelindex.org/models/llama-3/ · The Model Index · data as of 2026-10-01
