Skip to content
Model intelligence

34 AI models

Capability (Epoch Capabilities Index and individual benchmarks from Epoch AI), API list prices and context windows (OpenRouter), release dates and open-weight status. Benchmarks are shown individually - never merged into one invented score.
Leads on ECI
Claude Opus 5.5
ECI 167.35Epoch AI
Leads on GPQA Diamond
GPT-6 Astra
95.8%Epoch AI
Leads on SWE-bench Verified
Claude Opus 4.7
83.5%Epoch AI
Cheapest in ECI top 10
Claude Sonnet 5.5
$2.00/M inputOpenRouter
Longest context
Grok 4.20 Multi-Agent
2M tokensOpenRouter
Best open-weight on ECI
Kimi K3
ECI 157.61Epoch AI
Epoch Capabilities Index by release date

Capability frontier

Source: Epoch AI (CC BY 4.0). ECI is Epoch's item-response model over many benchmarks; each model page shows its 90% confidence interval.

Blended $/1M tokens (3:1 input:output) vs ECI

Price vs capability

Only models with both an ECI score and an OpenRouter price. Normalisation: 75% input + 25% output price. Click a point to open the model.

#ModelECI ↓ContextIn $/MReleased
1Muse Spark 1.3
Meta· reasoning
156.91.05M$1.252 Sep 2026
2Muse Spark 1.2
Meta· reasoning
155.01.05M$1.255 Aug 2026
3Muse Spark 1.1
Meta· reasoning
154.31.05M$1.259 Jul 2026
4Muse Spark152.1——8 Apr 2026
5Llama 4 Maverick132.21.05M$0.196 Apr 2025
6Llama 4 Scout129.71.31M$0.105 Apr 2025
7Llama 3.1-405B128.8——23 Jul 2024
8Llama 3.3 70B127.3——6 Dec 2024
9Llama 3.1-70B125.9——23 Jul 2024
10Llama 3.2 90B125.5——24 Sep 2024
11Llama 3-70B122.9——18 Apr 2024
12Llama 3.1-8B116.6——23 Jul 2024
13Llama 3-8B116.5——18 Apr 2024
14Llama 2-70B113.8——18 Jul 2023
15LLaMA-65B110.2——24 Feb 2023
16LLaMA-33B107.4——24 Feb 2023
17Llama 2-13B106.2——18 Jul 2023
18Llama 2-34B105.2——18 Jul 2023
19Llama 3.2 1B102.0——24 Sep 2024
20LLaMA-13B100.6——24 Feb 2023
21Llama 2-7B99.1——18 Jul 2023
22LLaMA-7B96.7——24 Feb 2023
23Muse Spark 1.3 Contributor
Meta· reasoning
—1.05M$0.102 Sep 2026
24Muse Spark 1.2 Contributor
Meta· reasoning
—1.05M$0.1021 Aug 2026
25Muse Glimmer 30B
Meta· reasoning
—131K$0.359 Aug 2026
26Llama Guard 4 12B—164K$0.1830 Apr 2025
27Llama 4 Maverick (FP8)———5 Apr 2025
28Llama 3.3 70B Instruct—131K$0.106 Dec 2024
29Llama 3.2 1B Instruct—60K$0.02725 Sep 2024
30Llama 3.2 3B Instruct—131K$0.05025 Sep 2024
31Llama 3.2 11B———24 Sep 2024
32Llama 3.2 3B———24 Sep 2024
33Llama 3.1 8B Instruct—131K$0.05023 Jul 2024
34Llama 3.1 70B Instruct—131K$0.4023 Jul 2024
Showing 34 of 34. Benchmark cells are % (best reported setting). 3.9K benchmark results in total.Export CSV