Skip to content
Model intelligence

10 AI models

Capability (Epoch Capabilities Index and individual benchmarks from Epoch AI), API list prices and context windows (OpenRouter), release dates and open-weight status. Benchmarks are shown individually - never merged into one invented score.
Leads on ECI
GPT-6 Astra
ECI 166.6Epoch AI
Leads on GPQA Diamond
GPT-6 Astra
95.8%Epoch AI
Leads on SWE-bench Verified
Claude Opus 4.7
83.5%Epoch AI
Cheapest in ECI top 10
GPT-5.6 Sol
$2.00/M inputOpenRouter
Longest context
Grok 4.20 Multi-Agent
2M tokensOpenRouter
Best open-weight on ECI
Kimi K3
ECI 157.68Epoch AI
Epoch Capabilities Index by release date

Capability frontier

Source: Epoch AI (CC BY 4.0). ECI is Epoch's item-response model over many benchmarks; each model page shows its 90% confidence interval.

Blended $/1M tokens (3:1 input:output) vs ECI

Price vs capability

Only models with both an ECI score and an OpenRouter price. Normalisation: 75% input + 25% output price. Click a point to open the model.

#ModelECI ↑ContextIn $/MReleased
1DeepSeek V4.1 Flash
DeepSeek· reasoning
155.01.05M$0.309 Sep 2026
2MiMo-V2.6-Pro
Xiaomi· reasoning
—1.05M$0.4321 Sep 2026
3MiMo-V2.6-Flash
Xiaomi· reasoning
—1.05M$0.1421 Sep 2026
4Ling 3.0 Flash VL—262K$0.02110 Sep 2026
5Granite 4.2 8B
IBM· reasoning
—131K$0.06031 Aug 2026
6Ternary Bonsai 2 27B
PrismML· reasoning
—262K$0.07518 Sep 2026
7Schematron V2 Turbo—128K$0.03012 Sep 2026
8Schematron V2 Small—128K$0.05012 Sep 2026
9Nex-N2.5-Pro
Nex AGI· reasoning
—262K$0.0758 Sep 2026
10Nex-N2.5-Mini
Nex AGI· reasoning
—262K$0.0258 Sep 2026
Showing 10 of 10. Benchmark cells are % (best reported setting). 3.8K benchmark results in total.Export CSV