> Qwen2 is an open-weights AI model from Alibaba, released 7 Jun 2024. Index score 103 (GPT-4 = 100), rank 180 of 229 ranked models.

[Models](https://themodelindex.org/models/) /[Alibaba](https://themodelindex.org/labs/alibaba/)

# Qwen2

Alibaba

Released 7 Jun 2024

Open

Date verified

[Compare](https://themodelindex.org/compare/qwen2-vs-claude-3-5-sonnet/) [Announcement](https://qwenlm.github.io/blog/qwen2/)

10391–114

Index score

#180 of 229 models

2/ 8

Core tests taken

2 of 6 areas · plus 4 anchor tests

−7

Versus the frontier at release

best then: GPT-4o

Replaced by [Qwen2.5](https://themodelindex.org/models/qwen2-5/)

–

Price per million tokens

no public API price found

–

Context window

tokens it reads at once

## Where it sits · on the frontier

_Chart: Model Index score of every scored model since late 2022, with the best closed and best open model over time_

GPT-4 Turbo reached this score 7 months earlier.

## Test results · by area · tick = best by any model

Show 4 anchor tests

Reasoning & knowledge

evidence

GPQA Diamond

41%

Humanity's Last Exam

no result

Maths

evidence

FrontierMath T1-3

no result

Coding

evidence

SWE-bench Verified

no result

Terminal-Bench

no result

Agents

evidence

APEX-Agents

no result

Novel problems

no evidence yet

ARC-AGI-2

no result

Long tasks

evidence

METR time horizon

2.2 min

Independent results from the Epoch AI benchmarking hub, best reasoning setting. The core tests are shown here; anchor tests also inform the score and can be shown above.

## What others say · rank on each leaderboard

The Model Index

independent tests, our method

#180

of 219 · 103

Epoch Capabilities Index

overall capability from many benchmarks

#153

of 239 · 125.3

LMArena · Maths

maths questions in chat

#252

of 373 · 1273

LMArena · Creative writing

creative writing

#285

of 386 · 1223

LMArena · Text

overall chat quality, judged by people

#288

of 388 · 1262

LMArena · Coding

coding questions in chat

#285

of 383 · 1297

Further left is better. Leaderboards measure different things: LMArena is people voting blind between two answers, the others are test scores. Click any row for the source.

## Best at · top three among current models

Replaced by Qwen2.5; it no longer leads any category among current models.

Where this model places in the top three of today's models, on category leaderboards and on our core tests.

## Facts and sources

Released

7 Jun 2024

Announcement

[qwenlm.github.io ↗](https://qwenlm.github.io/blog/qwen2/)

Release entry

Qwen2

Scored as

Qwen2-72B

Artificial Analysis index

–

Output speed

–

Hugging Face

not linked

## Alibaba releases · around this one

Qwen-72B

Open

30 NOV 2023

Qwen1.5

Open

4 FEB 2024

Qwen2

Open

7 JUN 2024

index 103

Qwen2.5

Open

19 SEP 2024

index 110

Qwen2.5-Coder-32B

Open

12 NOV 2024

---
Source: https://themodelindex.org/models/qwen2/ · The Model Index · data as of 2026-10-01
