The Model Index
Up to four models

Compare

Side by side on the same independent tests, with price and context. The link in your address bar shares this exact comparison.

Next to Grok 4.3 Beta
Presets
Grok 4.3 Beta
153136–170
WeightsClosed
ReleasedApr 2026
Price in / out–
Context–
Core tests taken2 of 9
Kimi K2.6
158141–175
WeightsOpen
ReleasedApr 2026
Price in / out$0.95 / $4.0
Context262K
Core tests taken4 of 9

Kimi K2.6 and Grok 4.3 Beta are too close to call (5 points apart, within the uncertainty), ahead on 2 of 2 shared core tests.

Test by testindependent results

Grok 4.3 BetaKimi K2.6
0%25%50%75%100%REASONING & KNOWLEDGEGPQA Diamond (Epoch AI)GPQA Diamond (Epoch AI)SimpleQA VerifiedSimpleQA VerifiedMATHSFrontierMath (tiers 1-3)FrontierMath (tiers 1-3)AIME-style maths (OTIS mock)AIME-style maths (OTIS mock)Chess puzzlesChess puzzlesFrontierMath (tier 4)FrontierMath (tier 4)CODINGWeirdMLWeirdMLAGENTSDTBenchDTBenchLMCALMCA

Tests at least two of these models have taken. Bold rows are the core tests.