Skip to content
Model intelligence

17 AI models

Capability (Epoch Capabilities Index and individual benchmarks from Epoch AI), API list prices and context windows (OpenRouter), release dates and open-weight status. Benchmarks are shown individually - never merged into one invented score.
Leads on ECI
GPT-6 Astra
ECI 166.6Epoch AI
Leads on GPQA Diamond
GPT-6 Astra
95.8%Epoch AI
Leads on SWE-bench Verified
Claude Opus 4.7
83.5%Epoch AI
Cheapest in ECI top 10
GPT-5.6 Sol
$2.00/M inputOpenRouter
Longest context
Grok 4.20 Multi-Agent
2M tokensOpenRouter
Best open-weight on ECI
Kimi K3
ECI 157.68Epoch AI
Epoch Capabilities Index by release date

Capability frontier

Source: Epoch AI (CC BY 4.0). ECI is Epoch's item-response model over many benchmarks; each model page shows its 90% confidence interval.

Blended $/1M tokens (3:1 input:output) vs ECI

Price vs capability

Only models with both an ECI score and an OpenRouter price. Normalisation: 75% input + 25% output price. Click a point to open the model.

#ModelECI ↓ContextIn $/MReleased
1Grok 4.6
xAI· reasoning
156.5500K$2.0012 Aug 2026
2Grok 4.5
xAI· reasoning
154.0500K$2.008 Jul 2026
3Grok 4.20
xAI· reasoning
152.02M$1.2517 Feb 2026
4Grok 4.3 Beta149.1——17 Apr 2026
5Grok 4146.5——9 Jul 2025
6Grok 4 Fast144.2——19 Sep 2025
7Grok-3 mini140.3——24 Jun 2025
8Grok 3138.3——9 Apr 2025
9Grok-2 (Dec 2024)130.5——12 Dec 2024
10Grok 4.7
xAI· reasoning
—500K$2.0021 Sep 2026
11Grok 4.3
xAI· reasoning
—1M$1.2530 Apr 2026
12Grok Build 0.1
xAI· reasoning
—256K$1.0020 May 2026
13Grok 4.20 Multi-Agent
xAI· reasoning
—2M$1.2531 Mar 2026
14Grok 4.1 Fast———19 Nov 2025
15Grok 4.1———17 Nov 2025
16grok-code-fast-1———28 Aug 2025
17Grok 4 Heavy (web app)———10 Jul 2025
Showing 17 of 17. Benchmark cells are % (best reported setting). 3.8K benchmark results in total.Export CSV