Skip to content
Model intelligence

676 AI models

Capability (Epoch Capabilities Index and individual benchmarks from Epoch AI), API list prices and context windows (OpenRouter), release dates and open-weight status. Benchmarks are shown individually - never merged into one invented score.
Leads on ECI
GPT-6 Astra
ECI 166.6Epoch AI
Leads on GPQA Diamond
GPT-6 Astra
95.8%Epoch AI
Leads on SWE-bench Verified
Claude Opus 4.7
83.5%Epoch AI
Cheapest in ECI top 10
GPT-5.6 Sol
$2.00/M inputOpenRouter
Longest context
Grok 4.20 Multi-Agent
2M tokensOpenRouter
Best open-weight on ECI
Kimi K3
ECI 157.68Epoch AI
Epoch Capabilities Index by release date

Capability frontier

Source: Epoch AI (CC BY 4.0). ECI is Epoch's item-response model over many benchmarks; each model page shows its 90% confidence interval.

Blended $/1M tokens (3:1 input:output) vs ECI

Price vs capability

Only models with both an ECI score and an OpenRouter price. Normalisation: 75% input + 25% output price. Click a point to open the model.

#ModelECI ↑ContextIn $/MReleased
101Claude 3.5 Haiku127.2——22 Oct 2024
102GPT-4 Turbo (Apr 2024)127.3——9 Apr 2024
103Llama 3.3 70B127.3——6 Dec 2024
104Mistral Small 3.1127.5——17 Mar 2025
105Mistral Large 2 (Jul 2024)127.5——24 Jul 2024
106Mistral Large 2 (Nov 2024)128.5——18 Nov 2024
107Qwen2.5-32B128.5——17 Sep 2024
108Llama 3.1-405B128.8——23 Jul 2024
109GPT-4o (Aug 2024)128.8——6 Aug 2024
110GPT-4o (Nov 2024)128.8——20 Nov 2024
111GPT-4o (May 2024)129.0——13 May 2024
112Qwen2.5-72B129.0——19 Sep 2024
113Gemini 1.5 Flash (Sep 2024)129.4——24 Sep 2024
114GPT-4.1 Nano129.61.05M$0.1014 Apr 2025
115Llama 4 Scout129.61.31M$0.105 Apr 2025
116Claude 3.5 Sonnet130.0——20 Jun 2024
117Gemma 3 27B130.0131K$0.08012 Mar 2025
118Phi 4130.416K$0.07012 Dec 2024
119Grok-2 (Dec 2024)130.5——12 Dec 2024
120Magistral Small 1.2131.4——18 Sep 2025
121Mistral Small 3.2131.7——20 Jun 2025
122Gemini 1.5 Pro (Sept 2024)131.7——24 Sep 2024
123Llama 4 Maverick132.21.05M$0.196 Apr 2025
124DeepSeek V3132.4164K$0.2626 Dec 2024
125Qwen2.5-Max132.5——25 Jan 2025
126Magistral Small 1.0133.2——10 Jun 2025
127Claude 3.5 Sonnet (October 2024)133.6——22 Oct 2024
128Gemini 2.5 Flash-Lite (Jun 2025)133.9——17 Jun 2025
129Mistral Medium 3134.1131K$0.407 May 2025
130Gemini 2.0 Flash (Feb 2025)134.7——5 Feb 2025
131Gemini 2.0 Flash (Dec 2024)134.7——11 Dec 2024
132o1-preview134.8——12 Sep 2024
133GPT-4.1 Mini135.01.05M$0.4014 Apr 2025
134Gemini 2.0 Pro135.1——5 Feb 2025
135Gemini 2.0 Flash Thinking (Jan 2025)135.4——21 Jan 2025
136DeepSeek-R1-Distill-Qwen-14B135.4——20 Jan 2025
137o1-mini135.8——12 Sep 2024
138DeepSeek-V3 (Mar 2025)135.9——24 Mar 2025
139Qwen3 8B
Alibaba (Qwen)· reasoning
136.2131K$0.1228 Apr 2025
140Qwen3 30B A3B
Alibaba (Qwen)· reasoning
136.2131K$0.1329 Apr 2025
141GPT-4.5136.7——27 Feb 2025
142GPT-4.1136.81.05M$2.0014 Apr 2025
143Qwen3-30B-A3B-Instruct (Jul 2025)137.4——29 Jul 2025
144DeepSeek-R1-Distill-Qwen-32B137.4——20 Jan 2025
145QwQ-32B137.6——5 Mar 2025
146gpt-oss-20b
OpenAI· reasoning
137.8131K$0.0185 Aug 2025
147Qwen3 14B
Alibaba (Qwen)· reasoning
138.2131K$0.1229 Apr 2025
148Grok 3138.3——9 Apr 2025
149Qwen3 32B
Alibaba (Qwen)· reasoning
138.5131K$0.08029 Apr 2025
150Qwen3-235B-A22B-Instruct (Jul 2025)138.9——25 Jul 2025
151R1
DeepSeek· reasoning
139.064K$0.7020 Jan 2025
152Qwen3 235B A22B
Alibaba (Qwen)· reasoning
139.4131K$0.4628 Apr 2025
153GPT-5 Nano
OpenAI· reasoning
139.4400K$0.0507 Aug 2025
154Qwen3.5-9B
Alibaba (Qwen)· reasoning
139.4262K$0.1024 Feb 2026
155Qwen3-30B-A3B-Thinking (Jul 2025)139.6——30 Jul 2025
156DeepSeek V3.1
DeepSeek· reasoning
139.9164K$0.2521 Aug 2025
157gpt-oss-120b
OpenAI· reasoning
139.9131K$0.0375 Aug 2025
158Gemini 2.5 Flash (Apr 2025)140.0——17 Apr 2025
159Kimi K2 (Jul 2025)140.1——12 Jul 2025
160Grok-3 mini140.3——24 Jun 2025
161o3 Mini
OpenAI· reasoning
140.3200K$1.1031 Jan 2025
162Gemini 2.5 Flash (Jun 2025)140.8——17 Jun 2025
163Claude 3.7 Sonnet141.2——24 Feb 2025
164DeepSeek-R1 (May 2025)141.3——28 May 2025
165Mistral Medium 3.5
Mistral AI· reasoning
141.4262K$1.5028 Apr 2026
166Gemini 2.5 Flash (May 2025)141.5——20 May 2025
167Claude Sonnet 4
Anthropic· reasoning
141.7200K$3.0022 May 2025
168Gemma 4 26B A4B
Google· reasoning
141.8262K$0.0902 Apr 2026
169o1
OpenAI· reasoning
141.9200K$15.017 Dec 2024
170Qwen3 Max
Alibaba (Qwen)· reasoning
142.4262K$0.7824 Sep 2025
171Claude Haiku 4.5
Anthropic· reasoning
142.4200K$1.0015 Oct 2025
172Gemini 2.5 Pro (May 2025)142.5——6 May 2025
173GPT-5.5 Instant142.5——5 May 2026
174Qwen3.5-35B-A3B
Alibaba (Qwen)· reasoning
142.5262K$0.1624 Feb 2026
175Claude Opus 4142.7——22 May 2025
176Gemma 4 31B
Google· reasoning
142.7262K$0.0902 Apr 2026
177Gemini 2.5 Flash (Sep 2025)143.0——25 Sep 2025
178Qwen 3.6 Flash143.3——27 Apr 2026
179GLM 4.7
Z.ai (Zhipu AI)· reasoning
143.5205K$0.6022 Dec 2025
180Qwen3-235B-A22B-Thinking (Jul 2025)143.9——25 Jul 2025
181Qwen 3.6 35B-A3B143.9——14 Apr 2026
182Qwen 3.5 Flash (hosted 35B-A3B)144.0——25 Feb 2026
183Claude Opus 4.1
Anthropic· reasoning
144.1200K$15.05 Aug 2025
184Gemini 2.5 Pro (Mar 2025)144.2——31 Mar 2025
185Grok 4 Fast144.2——19 Sep 2025
186Gemini 3.1 Flash Lite
Google· reasoning
144.41.05M$0.253 Mar 2026
187Qwen3.7 Flash
Alibaba (Qwen)· reasoning
144.61M$0.03027 Jul 2026
188DeepSeek V3.2 Exp
DeepSeek· reasoning
145.0164K$0.2729 Sep 2025
189Gemini 3.5 Flash Lite
Google· reasoning
145.11.05M$0.3021 Jul 2026
190Gemini 2.5 Pro (Jun 2025)145.3——5 Jun 2025
191GPT-5 Mini
OpenAI· reasoning
145.5400K$0.257 Aug 2025
192o4 Mini
OpenAI· reasoning
145.6200K$1.1016 Apr 2025
193GPT-5.4 Nano
OpenAI· reasoning
145.8400K$0.2017 Mar 2026
194MiniMax M2.7
MiniMax· reasoning
145.8205K$0.2118 Mar 2026
195GLM 5
Z.ai (Zhipu AI)· reasoning
145.8205K$0.6011 Feb 2026
196Kimi K2 Thinking
Moonshot AI· reasoning
146.0262K$0.606 Nov 2025
197DeepSeek V4 Flash 0423
DeepSeek· reasoning
146.11.05M$0.1424 Apr 2026
198Nemotron 3 Ultra
NVIDIA· reasoning
146.2262K$0.604 Jun 2026
199DeepSeek V3.2
DeepSeek· reasoning
146.3164K$0.281 Dec 2025
200Grok 4146.5——9 Jul 2025
Showing 100 of 676. Benchmark cells are % (best reported setting). 3.8K benchmark results in total.← PrevNext →Export CSV