Skip to content

Artificial Analysis index

Ranked by the Artificial Analysis Intelligence Index, as published via OpenRouter. Covers many new models before Epoch evaluates them.

Ranked by the Artificial Analysis Intelligence Index, as published via OpenRouter. Covers many new models before Epoch evaluates them. Source: Artificial Analysis (via OpenRouter). Variants (reasoning effort, batch) are folded into one row.

#ModelCapabilityReasoningCodingContext$/M
51DeepSeek V4 Flash 0423DeepSeek146.1——1.05M$0.17
52GPT-5.4 MiniOpenAI · +1 variant148.883.6—400K$1.69
53Nemotron 3 UltraNVIDIA146.285.4—262K$1.05
54MiniMax M2.7MiniMax145.8——205K$0.37
55Ling 3.0 Flash FininclusionAI (Ant Group)———262K$0.090
56Gemini 3.5 Flash LiteGoogle · +1 variant145.183.3—1.05M$0.85
57Qwen3.6 27BAlibaba (Qwen)146.585.9—262K$1.04
58Claude Sonnet 4.5Anthropic · +1 variant146.8——1M$6.00
59GPT-5.4 NanoOpenAI · +1 variant145.878.5—400K$0.46
60Ling 3.0 FlashinclusionAI (Ant Group)———262K$0.032
61LongCat 2.0Meituan———1.05M$0.53
62Qwen3.5 397B A17BAlibaba (Qwen)146.686.4—262K$1.29
63Qwen3.6 35B A3BAlibaba (Qwen)—84.8—262K$0.36
64Claude Haiku 4.5Anthropic · +1 variant142.471.2—200K$2.00
65GPT-5 MiniOpenAI · +1 variant145.575.0—400K$0.69
66Gemini 2.5 ProGoogle · +1 variant———1.05M$3.44
67Gemini 3.1 Flash Lite PreviewGoogle———1.05M$0.56
68Qwen3.5-122B-A10BAlibaba (Qwen)———262K$0.71
69DeepSeek V3.1 TerminusDeepSeek———164K$0.47
70Mistral Medium 3.5batchMistral AI · +1 variant———262K$1.50
71Command ACohere———256K$4.38
72Nemotron 3.5 LightningNVIDIA———262K$0.085
73Nemotron 3 SuperNVIDIA———262K$0.17
74Qwen3 235B A22B Thinking 2507Alibaba (Qwen)———131K$0.75
75Mercury 2.5NEWInception Labs———260K$0.068
76gpt-oss-120bOpenAI · +1 variant139.975.8—131K$0.070
77R1DeepSeek139.071.7—64K$1.15
78Mistral Small 4Mistral AI · +1 variant———262K$0.26
79Qwen3.5-9BAlibaba (Qwen)139.479.0—262K$0.11
80Granite 4.2 8BIBM———131K$0.11
81o3 Mini HighOpenAI———200K$1.93
82Trinity Large ThinkingArcee AI———262K$0.39
83Qwen3 30B A3B Thinking 2507Alibaba (Qwen)———82K$0.75
84DeepSeek V3 0324DeepSeek———164K$0.50
85Mistral Large 3 2512Mistral AI · +1 variant———262K$0.75
86Mistral Medium 3.1Mistral AI · +1 variant———131K$0.80
87Qwen3 Coder NextAlibaba (Qwen)———262K$0.29
88gpt-oss-20bOpenAI · +1 variant137.860.8—131K$0.036
89Nemotron 3 Nano 30B A3BNVIDIA———262K$0.087
90Devstral 2 2512Mistral AI———262K$0.80
91Solar Pro 3Upstage———131K$0.26
92Ministral 3 14B 2512Mistral AI———262K$0.20
93Ministral 3 8B 2512Mistral AI · +1 variant———262K$0.15
94Gemma 3 27BGoogle130.047.7—131K$0.17
95Ministral 3 3B 2512Mistral AI———131K$0.10
96Gemma 3 12BGoogle123.539.5—131K$0.075
— means no published score from that source yet · click a model for every benchmark with its source

New - awaiting independent evaluation

All new models →

Not ranked until an independent evaluator (Epoch AI) publishes scores. We don't use vendor-reported benchmark claims for rankings.

Best AI for coding →
Models ranked on DeepSWE
Best AI for reasoning →
Ranked on GPQA Diamond
Best AI for math →
Ranked on competition math
Best open-weight models →
Weights you can run yourself
Cheapest capable models →
Lowest blended API price
Best AI coding tools →
Apps & agents by popularity
Best AI image generators →
Apps & models
Best AI for research →
Answer engines & deep research