Skip to content

Newest models

Most recent releases, with whatever independent scores exist so far.

Most recent releases, with whatever independent scores exist so far. Source: OpenRouter / Epoch AI. Variants (reasoning effort, batch) are folded into one row.

#ModelCapabilityReasoningCodingContext$/M
501Qwen2-VL-72B-Instruct—————
502Qwen2-VL-7B-Instruct—————
503Llama 3.1 Euryale 70B v2.2Sao10K———131K$0.85
504Hermes 3 70B InstructNous Research———131K$0.70
505Phi-3.5-MoEMicrosoft—————
506Hermes 3 405B InstructNous Research———131K$1.00
507InternVL-Chat-ViT-6B-Vicuna-13B—————
508Phi-3.5-miniMicrosoft—————
509Phi-3.5-vision-instruct—————
510Llama 3 8B LunarisSao10K———8K$0.043
511GPT-4o (Aug 2024)OpenAI128.849.2———
512GPT-4o (2024-08-06)OpenAI———128K$4.38
513Mistral Large 2 (Jul 2024)Mistral AI127.549.0———
514Llama 3.1-405BMeta128.850.9———
515Llama 3.1-70BMeta125.944.2———
516Llama 3.1-8BMeta116.527.0———
517Llama 3.1 70B InstructMeta———131K$0.40
518Llama 3.1 8B InstructMeta———131K$0.058
519GPT-4o-miniOpenAI · +1 variant126.637.7—128K$0.26
520Mistral NemoMistral AI118.629.9—131K$0.022
521GPT-4o-mini (2024-07-18)OpenAI———128K$0.26
522Gemma 2 27BGoogle122.136.5—8K$0.65
523Gemma 2 9BGoogle119.827.5———
524Claude 3.5 SonnetAnthropic130.0————
525Claude 3.5 Sonnet (Jun 2024)Anthropic—54.0———
526Hermes 2 Theta Llama-3 70BNous Research—37.5———
527DeepSeek-Coder-V2 236BDeepSeek—————
528DeepSeek-Coder-V2-Lite-Base108.6————
529DeepSeek-Coder-V2-Lite-Instruct—————
530Qwen2-72BAlibaba (Qwen)125.340.8———
531Mistral 7B v0.3Mistral AI108.715.2———
532Gemini 1.5 Flash (May 2024)Google122.640.4———
533falcon-11B-vlm—————
534Gemini 1.5 Pro (May 2024)Google126.945.9———
535GPT-4o (May 2024)OpenAI129.048.9———
536GPT-4oOpenAI · +1 variant———128K$4.38
537GPT-4o (2024-05-13)OpenAI———128K$7.50
538Yi-1.5-34B01.AI—32.0———
539Yi-Large01.AI—————
540Falcon 2 11BTechnology Innovation Institute109.3————
541DeepSeek-V2 (MoE-236B, May 2024)DeepSeek124.8————
542phi-3-small 7.4BMicrosoft121.8————
543phi-3-medium 14BMicrosoft121.227.6———
544phi-3-mini 3.8BMicrosoft117.2————
545llama3-llava-next-8b—————
546Llama 3-70BMeta122.940.6———
547Llama 3-8BMeta116.326.1———
548Mixtral 8x22BMistral AI122.034.1———
549Mixtral 8x22B InstructMistral AI———66K$3.00
550WizardLM-2 8x22BMicrosoft—43.4—66K$0.62
— means no published score from that source yet · click a model for every benchmark with its source
Best AI for coding →
Models ranked on DeepSWE
Best AI for reasoning →
Ranked on GPQA Diamond
Best AI for math →
Ranked on competition math
Best open-weight models →
Weights you can run yourself
Cheapest capable models →
Lowest blended API price
Best AI coding tools →
Apps & agents by popularity
Best AI image generators →
Apps & models
Best AI for research →
Answer engines & deep research