Skip to content
Model intelligence

3 AI models

Capability (Epoch Capabilities Index and individual benchmarks from Epoch AI), API list prices and context windows (OpenRouter), release dates and open-weight status. Benchmarks are shown individually - never merged into one invented score.
Leads on ECI
GPT-6 Astra
ECI 166.6Epoch AI
Leads on GPQA Diamond
GPT-6 Astra
95.8%Epoch AI
Leads on SWE-bench Verified
Claude Opus 4.7
83.5%Epoch AI
Cheapest in ECI top 10
GPT-5.6 Sol
$2.00/M inputOpenRouter
Longest context
Grok 4.20 Multi-Agent
2M tokensOpenRouter
Best open-weight on ECI
Kimi K3
ECI 157.68Epoch AI
Epoch Capabilities Index by release date

Capability frontier

Source: Epoch AI (CC BY 4.0). ECI is Epoch's item-response model over many benchmarks; each model page shows its 90% confidence interval.

Blended $/1M tokens (3:1 input:output) vs ECI

Price vs capability

Only models with both an ECI score and an OpenRouter price. Normalisation: 75% input + 25% output price. Click a point to open the model.

#ModelECI ↑ContextIn $/MReleased
1Llama 3.3 Euryale 70B—131K$0.6518 Dec 2024
2Llama 3.1 Euryale 70B v2.2—131K$0.8528 Aug 2024
3Llama 3 8B Lunaris—8K$0.04013 Aug 2024
Showing 3 of 3. Benchmark cells are % (best reported setting). 3.8K benchmark results in total.Export CSV