> Llama 3.1-70B is an open-weights AI model from Meta, released 23 Jul 2024. Index score 104 (GPT-4 = 100), rank 178 of 229 ranked models. API price $0.40 in / $0.40 out per million tokens, 131K context.

[Models](https://themodelindex.org/models/) /[Meta](https://themodelindex.org/labs/meta/)

# Llama 3.1-70B

Meta

Released 23 Jul 2024

Open

[Compare](https://themodelindex.org/compare/llama-3-1-70b-vs-claude-3-5-sonnet/)

10492–115

Index score

#178 of 229 models

1/ 8

Core tests taken

1 of 6 areas · plus 8 anchor tests

−8

Versus the frontier at release

best then: Claude 3.5 Sonnet (Jun 2024)

Replaced by [Llama 3.3 70B](https://themodelindex.org/models/llama-3-3-70b/)

$0.40/ $0.40

Price per million tokens

input / output, list price

131K

Context window

tokens it reads at once

## Where it sits · on the frontier

_Chart: Model Index score of every scored model since late 2022, with the best closed and best open model over time_

GPT-4 Turbo reached this score 9 months earlier.

## Test results · by area · tick = best by any model

Show 8 anchor tests

Reasoning & knowledge

evidence

GPQA Diamond

44%

Humanity's Last Exam

no result

Maths

evidence

FrontierMath T1-3

no result

Coding

evidence

SWE-bench Verified

no result

Terminal-Bench

no result

Agents

evidence

APEX-Agents

no result

Novel problems

no evidence yet

ARC-AGI-2

no result

Long tasks

no evidence yet

METR time horizon

no result

Independent results from the Epoch AI benchmarking hub, best reasoning setting. The core tests are shown here; anchor tests also inform the score and can be shown above.

## What others say · rank on each leaderboard

The Model Index

independent tests, our method

#178

of 219 · 104

Epoch Capabilities Index

overall capability from many benchmarks

#150

of 239 · 125.9

LMArena · Text

overall chat quality, judged by people

#262

of 388 · 1293

LMArena · Coding

coding questions in chat

#259

of 383 · 1333

LMArena · Creative writing

creative writing

#265

of 386 · 1257

LMArena · Maths

maths questions in chat

#257

of 373 · 1269

Further left is better. Leaderboards measure different things: LMArena is people voting blind between two answers, the others are test scores. Click any row for the source.

## Best at · top three among current models

Replaced by Llama 3.3 70B; it no longer leads any category among current models.

Where this model places in the top three of today's models, on category leaderboards and on our core tests.

## Where to use it · prices checked 1 Oct 2026

#### API providers · 2 · per million tokens, in / out

D

DeepInfra

$0.40 / $0.40 · fp8

cheapest

A

Amazon Bedrock

$0.72 / $0.72

↗

#### Run it yourself

🤗

Official weights

meta-llama/Meta-Llama-3.1-70B-Instruct

↗

Provider prices come from OpenRouter's live listing; going direct to a provider can differ. Subscriptions are the lab's own apps.

## Facts and sources

Released

23 Jul 2024

Announcement

not linked yet

Release entry

–

Scored as

Llama 3.1-70B

Artificial Analysis index

–

Output speed

–

Hugging Face

[meta-llama/Meta-Llama-3.1-70B-Instruct ↗](https://huggingface.co/meta-llama/Meta-Llama-3.1-70B-Instruct)

## Meta releases · around this one

LLaMA

Open

24 FEB 2023

Llama 2

Open

18 JUL 2023

index 88

---
Source: https://themodelindex.org/models/llama-3-1-70b/ · The Model Index · data as of 2026-10-01
