Skip to content
Model intelligence

34 AI models

Capability (Epoch Capabilities Index and individual benchmarks from Epoch AI), API list prices and context windows (OpenRouter), release dates and open-weight status. Benchmarks are shown individually - never merged into one invented score.
Leads on ECI
GPT-6 Astra
ECI 166.6Epoch AI
Leads on GPQA Diamond
GPT-6 Astra
95.8%Epoch AI
Leads on SWE-bench Verified
Claude Opus 4.7
83.5%Epoch AI
Cheapest in ECI top 10
GPT-5.6 Sol
$2.00/M inputOpenRouter
Longest context
Grok 4.20 Multi-Agent
2M tokensOpenRouter
Best open-weight on ECI
Kimi K3
ECI 157.68Epoch AI
Epoch Capabilities Index by release date

Capability frontier

Source: Epoch AI (CC BY 4.0). ECI is Epoch's item-response model over many benchmarks; each model page shows its 90% confidence interval.

Blended $/1M tokens (3:1 input:output) vs ECI

Price vs capability

Only models with both an ECI score and an OpenRouter price. Normalisation: 75% input + 25% output price. Click a point to open the model.

#ModelECI ↑ContextIn $/MReleased
1LLaMA-7B96.1——24 Feb 2023
2Llama 2-7B98.5——18 Jul 2023
3LLaMA-13B100.1——24 Feb 2023
4Llama 3.2 1B102.0——24 Sep 2024
5Llama 2-34B104.8——18 Jul 2023
6Llama 2-13B105.8——18 Jul 2023
7LLaMA-33B107.1——24 Feb 2023
8LLaMA-65B109.9——24 Feb 2023
9Llama 2-70B113.6——18 Jul 2023
10Llama 3-8B116.3——18 Apr 2024
11Llama 3.1-8B116.5——23 Jul 2024
12Llama 3-70B122.9——18 Apr 2024
13Llama 3.2 90B125.5——24 Sep 2024
14Llama 3.1-70B125.9——23 Jul 2024
15Llama 3.3 70B127.3——6 Dec 2024
16Llama 3.1-405B128.8——23 Jul 2024
17Llama 4 Scout129.61.31M$0.105 Apr 2025
18Llama 4 Maverick132.21.05M$0.196 Apr 2025
19Muse Spark152.1——8 Apr 2026
20Muse Spark 1.1
Meta· reasoning
154.31.05M$1.259 Jul 2026
21Muse Spark 1.2
Meta· reasoning
155.21.05M$1.255 Aug 2026
22Muse Spark 1.3
Meta· reasoning
156.91.05M$1.252 Sep 2026
23Muse Spark 1.3 Contributor
Meta· reasoning
—1.05M$0.102 Sep 2026
24Muse Spark 1.2 Contributor
Meta· reasoning
—1.05M$0.1021 Aug 2026
25Muse Glimmer 30B
Meta· reasoning
—131K$0.309 Aug 2026
26Llama Guard 4 12B—164K$0.1830 Apr 2025
27Llama 4 Maverick (FP8)———5 Apr 2025
28Llama 3.3 70B Instruct—131K$0.106 Dec 2024
29Llama 3.2 1B Instruct—60K$0.02725 Sep 2024
30Llama 3.2 3B Instruct—131K$0.05025 Sep 2024
31Llama 3.2 11B———24 Sep 2024
32Llama 3.2 3B———24 Sep 2024
33Llama 3.1 8B Instruct—131K$0.05023 Jul 2024
34Llama 3.1 70B Instruct—131K$0.4023 Jul 2024
Showing 34 of 34. Benchmark cells are % (best reported setting). 3.8K benchmark results in total.Export CSV