Skip to content

Artificial Analysis index

Ranked by the Artificial Analysis Intelligence Index, as published via OpenRouter. Covers many new models before Epoch evaluates them.

Ranked by the Artificial Analysis Intelligence Index, as published via OpenRouter. Covers many new models before Epoch evaluates them. Source: Artificial Analysis (via OpenRouter). Variants (reasoning effort, batch) are folded into one row.

#ModelCapabilityReasoningCodingContext$/M
51GPT-5.4 MiniOpenAI · +1 variant148.883.6—400K$1.69
52Nemotron 3 UltraNVIDIA146.285.4—262K$1.05
53MiniMax M2.7MiniMax145.8——205K$0.37
54Ling 3.0 Flash FininclusionAI (Ant Group)———262K$0.090
55Gemini 3.5 Flash LiteGoogle · +1 variant145.183.3—1.05M$0.85
56Qwen3.6 27BAlibaba (Qwen)146.585.9—262K$1.04
57Claude Sonnet 4.5Anthropic · +1 variant146.8——1M$6.00
58GPT-5.4 NanoOpenAI · +1 variant145.878.5—400K$0.46
59LongCat 2.0Meituan———1.05M$0.53
60Qwen3.5 397B A17BAlibaba (Qwen)146.686.4—262K$1.29
61Qwen3.6 35B A3BAlibaba (Qwen)—84.8—262K$0.36
62Claude Haiku 4.5Anthropic · +1 variant142.471.2—200K$2.00
63GPT-5 MiniOpenAI · +1 variant145.575.0—400K$0.69
64Gemini 2.5 ProGoogle · +1 variant———1.05M$3.44
65Gemini 3.1 Flash Lite PreviewGoogle———1.05M$0.56
66Qwen3.5-122B-A10BAlibaba (Qwen)———262K$0.71
67DeepSeek V3.1 TerminusDeepSeek———164K$0.47
68Mistral Medium 3.5batchMistral AI · +1 variant———262K$1.50
69Command ACohere———256K$4.38
70Nemotron 3.5 LightningNVIDIA———1M$0.085
71Nemotron 3 SuperNVIDIA———262K$0.17
72Qwen3 235B A22B Thinking 2507Alibaba (Qwen)———131K$0.75
73Mercury 2.5NEWInception Labs———260K$0.068
74gpt-oss-120bOpenAI · +1 variant139.975.8—131K$0.070
75R1DeepSeek139.071.7—64K$1.15
76Mistral Small 4Mistral AI · +1 variant———262K$0.26
77Qwen3.5-9BAlibaba (Qwen)139.479.0—262K$0.11
78Granite 4.2 8BNEWIBM———131K$0.11
79o3 Mini HighOpenAI———200K$1.93
80Trinity Large ThinkingArcee AI———262K$0.39
81Qwen3 30B A3B Thinking 2507Alibaba (Qwen)———82K$0.75
82DeepSeek V3 0324DeepSeek———164K$0.50
83Mistral Large 3 2512Mistral AI · +1 variant———262K$0.75
84Mistral Medium 3.1Mistral AI · +1 variant———131K$0.80
85Qwen3 Coder NextAlibaba (Qwen)———262K$0.29
86gpt-oss-20bOpenAI · +1 variant137.860.8—131K$0.036
87Nemotron 3 Nano 30B A3BNVIDIA———262K$0.087
88Devstral 2 2512Mistral AI———262K$0.80
89Solar Pro 3Upstage———131K$0.26
90Ministral 3 14B 2512Mistral AI———262K$0.20
91Ministral 3 8B 2512Mistral AI · +1 variant———262K$0.15
92Gemma 3 27BGoogle130.047.7—131K$0.17
93Ministral 3 3B 2512Mistral AI———131K$0.10
94Gemma 3 12BGoogle123.539.5—131K$0.075
— means no published score from that source yet · click a model for every benchmark with its source

New - awaiting independent evaluation

All new models →

Not ranked until an independent evaluator (Epoch AI) publishes scores. We don't use vendor-reported benchmark claims for rankings.

Best AI for coding →
Models ranked on DeepSWE
Best AI for reasoning →
Ranked on GPQA Diamond
Best AI for math →
Ranked on competition math
Best open-weight models →
Weights you can run yourself
Cheapest capable models →
Lowest blended API price
Best AI coding tools →
Apps & agents by popularity
Best AI image generators →
Apps & models
Best AI for research →
Answer engines & deep research