Model intelligence
7 AI models
Capability (Epoch Capabilities Index and individual benchmarks from Epoch AI), API list prices and context windows (OpenRouter), release dates and open-weight status. Benchmarks are shown individually - never merged into one invented score.
Epoch Capabilities Index by release date
Capability frontier
Source: Epoch AI (CC BY 4.0). ECI is Epoch's item-response model over many benchmarks; each model page shows its 90% confidence interval.
Blended $/1M tokens (3:1 input:output) vs ECI
Price vs capability
Only models with both an ECI score and an OpenRouter price. Normalisation: 75% input + 25% output price. Click a point to open the model.
| # | Model | ECI ↑ | Context | In $/M | Weights | Released |
|---|---|---|---|---|---|---|
| 1 | Grok 4.20 xAI· reasoning | 152.0 | 2M | $1.25 | Closed | 17 Feb 2026 |
| 2 | Grok 4.5 xAI· reasoning | 154.0 | 500K | $2.00 | Closed | 8 Jul 2026 |
| 3 | Grok 4.6 xAI· reasoning | 156.5 | 500K | $2.00 | Closed | 12 Aug 2026 |
| 4 | Grok 4.7 xAI· reasoning | — | 500K | $2.00 | ? | 21 Sep 2026 |
| 5 | Grok 4.3 xAI· reasoning | — | 1M | $1.25 | ? | 30 Apr 2026 |
| 6 | Grok Build 0.1 xAI· reasoning | — | 256K | $1.00 | ? | 20 May 2026 |
| 7 | Grok 4.20 Multi-Agent xAI· reasoning | — | 2M | $1.25 | ? | 31 Mar 2026 |
Showing 7 of 7. Benchmark cells are % (best reported setting). 3.8K benchmark results in total.Export CSV