> GLM-5.3-Flash is an open-weights AI model from Zhipu, released 20 Aug 2026. Index score 158 (GPT-4 = 100), rank 45 of 229 ranked models. API price $0.15 in / $0.50 out per million tokens, 1.05M context.

[Models](https://themodelindex.org/models/) /[Zhipu](https://themodelindex.org/labs/zhipu/)

# GLM-5.3-Flash

Zhipu

Released 20 Aug 2026

Open

[Compare](https://themodelindex.org/compare/glm-5-3-flash-vs-gpt-6-astra/) [Run it yourself →](https://themodelindex.org/local/) [Announcement](https://huggingface.co/zai-org/GLM-5.3-Flash)

158146–169

Index score

#45 of 229 models

3/ 8

Core tests taken

3 of 6 areas · plus 8 anchor tests

−29

Versus the frontier at release

best then: Claude Opus 5

$0.15/ $0.50

Price per million tokens

input / output, list price

1.05M

Context window

tokens it reads at once

## Where it sits · on the frontier

_Chart: Model Index score of every scored model since late 2022, with the best closed and best open model over time_

Gemini 3 Pro reached this score 9 months earlier.

## Test results · by area · tick = best by any model

Show 7 anchor tests

Reasoning & knowledge

evidence

GPQA Diamond

90%

Humanity's Last Exam

no result

Maths

evidence

FrontierMath T1-3

56%

Coding

evidence

SWE-bench Verified

no result

Terminal-Bench

no result

Agents

evidence

APEX-Agents

53%

Novel problems

no evidence yet

ARC-AGI-2

no result

Long tasks

no evidence yet

METR time horizon

no result

Independent results from the Epoch AI benchmarking hub, best reasoning setting. The core tests are shown here; anchor tests also inform the score and can be shown above.

## What others say · rank on each leaderboard

The Model Index

independent tests, our method

#45

of 219 · 158

LMArena · Maths

maths questions in chat

#14

of 373 · 1500

LMArena · Coding

coding questions in chat

#25

of 383 · 1523

Artificial Analysis Intelligence Index

overall capability across ten evaluations

#16

of 207 · 42

LMArena · Text

overall chat quality, judged by people

#35

of 388 · 1474

LMArena · Vision

understanding images

#21

of 149 · 1281

LMArena · Creative writing

creative writing

#59

of 386 · 1435

Epoch Capabilities Index

overall capability from many benchmarks

#42

of 239 · 151.9

LMArena · WebDev

building web apps and front ends

#22

of 123 · 1615

Further left is better. Leaderboards measure different things: LMArena is people voting blind between two answers, the others are test scores. Click any row for the source.

## Best at · top three among current models

Not in the top three among current models on any category we track.

Where this model places in the top three of today's models, on category leaderboards and on our core tests.

## Where to use it · prices checked 1 Oct 2026

#### From the lab

Z.ai Chat

app · subscription

↗

Z.ai API

API

↗

#### API providers · 31 · per million tokens, in / out

O

OpenInference

$0.020 / $0.30 · fp4

cheapest

D

DeepInfra

$0.075 / $0.25 · fp4

↗

N

Novita

$0.084 / $0.28 · fp8

↗

S

StreamLake

$0.087 / $0.29 · fp8

↗

G

GMICloud

$0.090 / $0.30 · fp8

↗

R

Relace

$0.035 / $0.50

↗

S

Sail Research

$0.045 / $0.60 · fp4

↗

I

InferenceNet

$0.050 / $0.60 · fp4

↗

#### Run it yourself

🤗

Official weights

zai-org/GLM-5.3-Flash

↗

Provider prices come from OpenRouter's live listing; going direct to a provider can differ. Subscriptions are the lab's own apps.

## Facts and sources

Released

20 Aug 2026

Announcement

[huggingface.co ↗](https://huggingface.co/zai-org/GLM-5.3-Flash)

Release entry

GLM-5.3-Flash

Scored as

GLM-5.3-Flash

Artificial Analysis index

42

Output speed

45 tokens/s

Hugging Face

[zai-org/GLM-5.3-Flash ↗](https://huggingface.co/zai-org/GLM-5.3-Flash)

## Zhipu releases · around this one

GLM-5.2

Open

16 JUN 2026

index 158

GLM-5.3

Open

14 AUG 2026

index 167

GLM-5.3-Flash

Open

20 AUG 2026

index 158

---
Source: https://themodelindex.org/models/glm-5-3-flash/ · The Model Index · data as of 2026-10-01
