Skip to content
Model intelligence

291 AI models

Capability (Epoch Capabilities Index and individual benchmarks from Epoch AI), API list prices and context windows (OpenRouter), release dates and open-weight status. Benchmarks are shown individually - never merged into one invented score.
Leads on ECI
GPT-6 Astra
ECI 166.6Epoch AI
Leads on GPQA Diamond
GPT-6 Astra
95.8%Epoch AI
Leads on SWE-bench Verified
Claude Opus 4.7
83.5%Epoch AI
Cheapest in ECI top 10
GPT-5.6 Sol
$2.00/M inputOpenRouter
Longest context
Grok 4.20 Multi-Agent
2M tokensOpenRouter
Best open-weight on ECI
Kimi K3
ECI 157.68Epoch AI
Epoch Capabilities Index by release date

Capability frontier

Source: Epoch AI (CC BY 4.0). ECI is Epoch's item-response model over many benchmarks; each model page shows its 90% confidence interval.

Blended $/1M tokens (3:1 input:output) vs ECI

Price vs capability

Only models with both an ECI score and an OpenRouter price. Normalisation: 75% input + 25% output price. Click a point to open the model.

#ModelECI ↑ContextIn $/MReleased
1DeepSeek Coder 1.3B62.2——2 Nov 2023
2Cerebras-GPT-13B82.4——20 Mar 2023
3StarCoder 2 3B88.0——22 Feb 2024
4DeepSeek Coder 6.7B88.9——2 Nov 2023
5Dolly 2.0-12b
Databricks
88.9——11 Apr 2023
6Baichuan1-7B
Baichuan
89.8——1 Jun 2023
7Phi-1.590.8——11 Sep 2023
8XGen-7B
Salesforce
92.8——27 Jun 2023
9StarCoder 2 7B92.9——20 Feb 2024
10Gemma 2B93.6——21 Feb 2024
11MPT-7B
MosaicML
94.0——5 May 2023
12Falcon-7B94.5——24 Apr 2023
13DeepSeek Coder 33B95.7——2 Nov 2023
14Baichuan 2-7B
Baichuan
95.8——20 Sep 2023
15LLaMA-7B96.1——24 Feb 2023
16Llama 2-7B98.5——18 Jul 2023
17LLaMA-13B100.1——24 Feb 2023
18MPT-30B
MosaicML
100.2——22 Jun 2023
19Llama 3.2 1B102.0——24 Sep 2024
20Qwen2.5-Coder (1.5B)102.6——18 Sep 2024
21Baichuan2-13B
Baichuan
102.8——6 Sep 2023
22Falcon-40B104.0——25 May 2023
23Yi 6B
01.AI
104.4——22 Nov 2023
24StarCoder 2 15B104.6——20 Feb 2024
25Llama 2-13B105.8——18 Jul 2023
26Qwen-7B106.5——28 Sep 2023
27LLaMA-33B107.1——24 Feb 2023
28Phi-2107.6——12 Dec 2023
29Mistral 7B v0.3108.7——27 May 2024
30Falcon 2 11B109.3——9 May 2024
31LLaMA-65B109.9——24 Feb 2023
32DeepSeek LLM 67B110.5——29 Nov 2023
33Gemma 7B111.7——21 Feb 2024
34Falcon-180B111.9——6 Sep 2023
35Mistral 7B v0.1112.0——27 Sep 2023
36Qwen-14B112.8——24 Sep 2023
37Qwen2.5-Coder (7B)112.9——18 Sep 2024
38Llama 2-70B113.6——18 Jul 2023
39Gemma 3 4B116.0131K$0.05012 Mar 2025
40Llama 3-8B116.3——18 Apr 2024
41Llama 3.1-8B116.5——23 Jul 2024
42Stable Beluga 2117.0——20 Jul 2023
43phi-3-mini 3.8B117.2——23 Apr 2024
44Yi-34B
01.AI
117.3——2 Nov 2023
45Mixtral 8x7B118.4——11 Dec 2023
46Qwen2.5-7B118.4——19 Sep 2024
47Mistral Nemo118.6131K$0.01918 Jul 2024
48Command R+119.3——30 Aug 2024
49Qwen2.5-Coder-32B119.4——18 Sep 2024
50Gemma 2 9B119.8——24 Jun 2024
51phi-3-medium 14B121.2——23 Apr 2024
52phi-3-small 7.4B121.8——23 Apr 2024
53Mixtral 8x22B122.0——17 Apr 2024
54Gemma 2 27B122.18K$0.6524 Jun 2024
55Llama 3-70B122.9——18 Apr 2024
56Gemma 3 12B123.5131K$0.05012 Mar 2025
57DeepSeek-V2 (MoE-236B, May 2024)124.8——7 May 2024
58Qwen2-72B125.3——7 Jun 2024
59Llama 3.2 90B125.5——24 Sep 2024
60Llama 3.1-70B125.9——23 Jul 2024
61Mistral Small 3127.133K$0.05030 Jan 2025
62Llama 3.3 70B127.3——6 Dec 2024
63Mistral Small 3.1127.5——17 Mar 2025
64Mistral Large 2 (Jul 2024)127.5——24 Jul 2024
65Mistral Large 2 (Nov 2024)128.5——18 Nov 2024
66Qwen2.5-32B128.5——17 Sep 2024
67Llama 3.1-405B128.8——23 Jul 2024
68Qwen2.5-72B129.0——19 Sep 2024
69Llama 4 Scout129.61.31M$0.105 Apr 2025
70Gemma 3 27B130.0131K$0.08012 Mar 2025
71Phi 4130.416K$0.07012 Dec 2024
72Magistral Small 1.2131.4——18 Sep 2025
73Mistral Small 3.2131.7——20 Jun 2025
74Llama 4 Maverick132.21.05M$0.196 Apr 2025
75DeepSeek V3132.4164K$0.2626 Dec 2024
76Magistral Small 1.0133.2——10 Jun 2025
77DeepSeek-R1-Distill-Qwen-14B135.4——20 Jan 2025
78DeepSeek-V3 (Mar 2025)135.9——24 Mar 2025
79Qwen3 8B
Alibaba (Qwen)· reasoning
136.2131K$0.1228 Apr 2025
80Qwen3 30B A3B
Alibaba (Qwen)· reasoning
136.2131K$0.1329 Apr 2025
81Qwen3-30B-A3B-Instruct (Jul 2025)137.4——29 Jul 2025
82DeepSeek-R1-Distill-Qwen-32B137.4——20 Jan 2025
83QwQ-32B137.6——5 Mar 2025
84gpt-oss-20b
OpenAI· reasoning
137.8131K$0.0185 Aug 2025
85Qwen3 14B
Alibaba (Qwen)· reasoning
138.2131K$0.1229 Apr 2025
86Qwen3 32B
Alibaba (Qwen)· reasoning
138.5131K$0.08029 Apr 2025
87Qwen3-235B-A22B-Instruct (Jul 2025)138.9——25 Jul 2025
88R1
DeepSeek· reasoning
139.064K$0.7020 Jan 2025
89Qwen3 235B A22B
Alibaba (Qwen)· reasoning
139.4131K$0.4628 Apr 2025
90Qwen3.5-9B
Alibaba (Qwen)· reasoning
139.4262K$0.1024 Feb 2026
91Qwen3-30B-A3B-Thinking (Jul 2025)139.6——30 Jul 2025
92DeepSeek V3.1
DeepSeek· reasoning
139.9164K$0.2521 Aug 2025
93gpt-oss-120b
OpenAI· reasoning
139.9131K$0.0375 Aug 2025
94Kimi K2 (Jul 2025)140.1——12 Jul 2025
95DeepSeek-R1 (May 2025)141.3——28 May 2025
96Mistral Medium 3.5
Mistral AI· reasoning
141.4262K$1.5028 Apr 2026
97Gemma 4 26B A4B
Google· reasoning
141.8262K$0.0902 Apr 2026
98Qwen3.5-35B-A3B
Alibaba (Qwen)· reasoning
142.5262K$0.1624 Feb 2026
99Gemma 4 31B
Google· reasoning
142.7262K$0.0902 Apr 2026
100GLM 4.7
Z.ai (Zhipu AI)· reasoning
143.5205K$0.6022 Dec 2025
Showing 100 of 291. Benchmark cells are % (best reported setting). 3.8K benchmark results in total.Next →Export CSV