Skip to content

Newest models

Most recent releases, with whatever independent scores exist so far.

Most recent releases, with whatever independent scores exist so far. Source: OpenRouter / Epoch AI. Variants (reasoning effort, batch) are folded into one row.

#ModelCapabilityReasoningCodingContext$/M
501c4ai-command-r-08-2024—————
502Command R (08-2024)Cohere———128K$0.26
503Qwen2-VL-72B-Instruct—————
504Qwen2-VL-7B-Instruct—————
505Llama 3.1 Euryale 70B v2.2Sao10K———131K$0.85
506Hermes 3 70B InstructNous Research———131K$0.70
507Phi-3.5-MoEMicrosoft—————
508Hermes 3 405B InstructNous Research———131K$1.00
509InternVL-Chat-ViT-6B-Vicuna-13B—————
510Phi-3.5-miniMicrosoft—————
511Phi-3.5-vision-instruct—————
512Llama 3 8B LunarisSao10K———8K$0.043
513GPT-4o (Aug 2024)OpenAI128.849.2———
514GPT-4o (2024-08-06)OpenAI———128K$4.38
515Mistral Large 2 (Jul 2024)Mistral AI127.549.0———
516Llama 3.1-405BMeta128.850.9———
517Llama 3.1-70BMeta125.944.2———
518Llama 3.1-8BMeta116.527.0———
519Llama 3.1 70B InstructMeta———131K$0.40
520Llama 3.1 8B InstructMeta———131K$0.058
521GPT-4o-miniOpenAI · +1 variant126.637.7—128K$0.26
522Mistral NemoMistral AI118.629.9—131K$0.022
523GPT-4o-mini (2024-07-18)OpenAI———128K$0.26
524Gemma 2 27BGoogle122.136.5—8K$0.65
525Gemma 2 9BGoogle119.827.5———
526Claude 3.5 SonnetAnthropic130.0————
527Claude 3.5 Sonnet (Jun 2024)Anthropic—54.0———
528Hermes 2 Theta Llama-3 70BNous Research—37.5———
529DeepSeek-Coder-V2 236BDeepSeek—————
530DeepSeek-Coder-V2-Lite-Base108.6————
531DeepSeek-Coder-V2-Lite-Instruct—————
532Qwen2-72BAlibaba (Qwen)125.340.8———
533Mistral 7B v0.3Mistral AI108.715.2———
534Gemini 1.5 Flash (May 2024)Google122.640.4———
535falcon-11B-vlm—————
536Gemini 1.5 Pro (May 2024)Google126.945.9———
537GPT-4o (May 2024)OpenAI129.048.9———
538GPT-4oOpenAI · +1 variant———128K$4.38
539GPT-4o (2024-05-13)OpenAI———128K$7.50
540Yi-1.5-34B01.AI—32.0———
541Yi-Large01.AI—————
542Falcon 2 11BTechnology Innovation Institute109.3————
543DeepSeek-V2 (MoE-236B, May 2024)DeepSeek124.8————
544phi-3-small 7.4BMicrosoft121.8————
545phi-3-medium 14BMicrosoft121.227.6———
546phi-3-mini 3.8BMicrosoft117.2————
547llama3-llava-next-8b—————
548Llama 3-70BMeta122.940.6———
549Llama 3-8BMeta116.326.1———
550Mixtral 8x22BMistral AI122.034.1———
— means no published score from that source yet · click a model for every benchmark with its source
Best AI for coding →
Models ranked on DeepSWE
Best AI for reasoning →
Ranked on GPQA Diamond
Best AI for math →
Ranked on competition math
Best open-weight models →
Weights you can run yourself
Cheapest capable models →
Lowest blended API price
Best AI coding tools →
Apps & agents by popularity
Best AI image generators →
Apps & models
Best AI for research →
Answer engines & deep research