> o3-mini is a closed AI model from OpenAI, released 31 Jan 2025. Index score 132 (GPT-4 = 100), rank 109 of 229 ranked models. API price $1.1 in / $4.4 out per million tokens, 200K context.

[Models](https://themodelindex.org/models/) /[OpenAI](https://themodelindex.org/labs/openai/)

# o3-mini

OpenAI

Released 31 Jan 2025

Closed

Reasoning

Date verified

[Compare](https://themodelindex.org/compare/o3-mini-2025-01-vs-deepseek-r1/) [Announcement](https://openai.com/index/openai-o3-mini/)

132121–144

Index score

#109 of 229 models

3/ 8

Core tests taken

3 of 6 areas · plus 17 anchor tests

−3

Versus the frontier at release

best then: o1

Replaced by [o4-mini](https://themodelindex.org/models/o4-mini/)

$1.1/ $4.4

Price per million tokens

input / output, list price

200K

Context window

tokens it reads at once

## Where it sits · on the frontier

_Chart: Model Index score of every scored model since late 2022, with the best closed and best open model over time_

o1 reached this score 2 months earlier.

## Test results · by area · tick = best by any model

Show 13 anchor tests

Reasoning & knowledge

evidence

GPQA Diamond

77%

Humanity's Last Exam

no result

Maths

evidence

FrontierMath T1-3

19%

Coding

evidence

SWE-bench Verified

no result

Terminal-Bench

no result

Agents

evidence

APEX-Agents

no result

Novel problems

evidence

ARC-AGI-2

3%

Long tasks

no evidence yet

METR time horizon

no result

Independent results from the Epoch AI benchmarking hub, best reasoning setting. The core tests are shown here; anchor tests also inform the score and can be shown above.

## What others say · rank on each leaderboard

The Model Index

independent tests, our method

#109

of 219 · 132

LMArena · Maths

maths questions in chat

#130

of 373 · 1405

LMArena · Coding

coding questions in chat

#155

of 383 · 1434

Epoch Capabilities Index

overall capability from many benchmarks

#103

of 239 · 140.3

LMArena · Text

overall chat quality, judged by people

#179

of 388 · 1364

LMArena · Creative writing

creative writing

#194

of 386 · 1312

Further left is better. Leaderboards measure different things: LMArena is people voting blind between two answers, the others are test scores. Click any row for the source.

## Best at · top three among current models

Replaced by o4-mini; it no longer leads any category among current models.

Where this model places in the top three of today's models, on category leaderboards and on our core tests.

## Where to use it · prices checked 1 Oct 2026

#### From the lab

ChatGPT

app · subscription

↗

OpenAI API

API

↗

#### API providers · 1 · per million tokens, in / out

O

OpenAI

$1.1 / $4.4

cheapest

Provider prices come from OpenRouter's live listing; going direct to a provider can differ. Subscriptions are the lab's own apps.

## Facts and sources

Released

31 Jan 2025

Announcement

[openai.com ↗](https://openai.com/index/openai-o3-mini/)

Release entry

o3-mini

Scored as

o3-mini

Artificial Analysis index

–

Output speed

–

Hugging Face

closed weights

## OpenAI releases · around this one

o1-preview

Closed

12 SEP 2024

index 121

o1

Closed

5 DEC 2024

index 136

o3-mini

Closed

31 JAN 2025

index 132

GPT-4.5

Closed

27 FEB 2025

index 124

GPT-4.1

Closed

14 APR 2025

index 125

---
Source: https://themodelindex.org/models/o3-mini-2025-01/ · The Model Index · data as of 2026-10-01
