Skip to content

Open-weight leaderboard

Models whose weights are published, ranked by the Epoch Capabilities Index.

Models whose weights are published, ranked by the Epoch Capabilities Index. Source: Epoch AI. Variants (reasoning effort, batch) are folded into one row.

#ModelCapabilityReasoningCodingContext$/M
1Kimi K3Moonshot AI157.793.168.51.05M$6.00
2GLM 5.3Z.ai (Zhipu AI)155.690.969.01.31M$2.15
3DeepSeek V4 Pro 0813DeepSeek155.491.7—1.05M$1.41
4DeepSeek V4.1 FlashNEWDeepSeek155.0——1.05M$0.53
5DeepSeek V4 Flash 0731DeepSeek154.591.0—1.31M$0.093
6GLM 5.3 FlashZ.ai (Zhipu AI)151.990.263.41.31M$0.24
7Inkling SmallThinking Machines Lab150.1——524K$0.64
8Qwen 3.8 27BAlibaba (Qwen)149.4————
9InklingThinking Machines Lab148.6——524K$1.76
— means no published score from that source yet · click a model for every benchmark with its source

New - awaiting independent evaluation

All new models →

Not ranked until an independent evaluator (Epoch AI) publishes scores. We don't use vendor-reported benchmark claims for rankings.

Best AI for coding →
Models ranked on DeepSWE
Best AI for reasoning →
Ranked on GPQA Diamond
Best AI for math →
Ranked on competition math
Best open-weight models →
Weights you can run yourself
Cheapest capable models →
Lowest blended API price
Best AI coding tools →
Apps & agents by popularity
Best AI image generators →
Apps & models
Best AI for research →
Answer engines & deep research