> Llama 2 is an open-weights AI model from Meta, released 18 Jul 2023. Index score 88 (GPT-4 = 100), rank 208 of 229 ranked models.

[Models](https://themodelindex.org/models/) /[Meta](https://themodelindex.org/labs/meta/)

# Llama 2

Meta

Released 18 Jul 2023

Open

Date verified

[Compare](https://themodelindex.org/compare/llama-2-vs-gpt-4/) [Announcement](https://ai.meta.com/research/publications/llama-2-open-foundation-and-fine-tuned-chat-models/)

8876–101

Index score

#208 of 229 models

1/ 8

Core tests taken

1 of 6 areas · plus 5 anchor tests

−12

Versus the frontier at release

best then: GPT-4

Replaced by [Llama 3](https://themodelindex.org/models/llama-3/)

–

Price per million tokens

no public API price found

–

Context window

tokens it reads at once

## Where it sits · on the frontier

_Chart: Model Index score of every scored model since late 2022, with the best closed and best open model over time_

GPT-4 reached this score 4 months earlier.

## Test results · by area · tick = best by any model

Show 5 anchor tests

Reasoning & knowledge

evidence

GPQA Diamond

26%

Humanity's Last Exam

no result

Maths

evidence

FrontierMath T1-3

no result

Coding

no evidence yet

SWE-bench Verified

no result

Terminal-Bench

no result

Agents

evidence

APEX-Agents

no result

Novel problems

no evidence yet

ARC-AGI-2

no result

Long tasks

no evidence yet

METR time horizon

no result

Independent results from the Epoch AI benchmarking hub, best reasoning setting. The core tests are shown here; anchor tests also inform the score and can be shown above.

## What others say · rank on each leaderboard

The Model Index

independent tests, our method

#208

of 219 · 88

Epoch Capabilities Index

overall capability from many benchmarks

#185

of 239 · 113.8

LMArena · Text

overall chat quality, judged by people

#340

of 388 · 1171

LMArena · Maths

maths questions in chat

#332

of 373 · 1137

LMArena · Coding

coding questions in chat

#347

of 383 · 1179

LMArena · Creative writing

creative writing

#350

of 386 · 1112

Further left is better. Leaderboards measure different things: LMArena is people voting blind between two answers, the others are test scores. Click any row for the source.

## Best at · top three among current models

Replaced by Llama 3; it no longer leads any category among current models.

Where this model places in the top three of today's models, on category leaderboards and on our core tests.

## Facts and sources

Released

18 Jul 2023

Announcement

[ai.meta.com ↗](https://ai.meta.com/research/publications/llama-2-open-foundation-and-fine-tuned-chat-models/)

Release entry

Llama 2

Scored as

Llama 2-70B

Artificial Analysis index

–

Output speed

–

Hugging Face

not linked

## Meta releases · around this one

LLaMA

Open

24 FEB 2023

Llama 2

Open

18 JUL 2023

index 88

Code Llama

Open

24 AUG 2023

Llama 3

Open

18 APR 2024

index 98

---
Source: https://themodelindex.org/models/llama-2/ · The Model Index · data as of 2026-10-01
