Skip to content

Newest models

Most recent releases, with whatever independent scores exist so far.

Most recent releases, with whatever independent scores exist so far. Source: OpenRouter / Epoch AI. Variants (reasoning effort, batch) are folded into one row.

#ModelCapabilityReasoningCodingContext$/M
551CodeQwen1.5-7B94.2————
552GPT-4 Turbo (Apr 2024)OpenAI127.346.6———
553GPT-4 TurboOpenAI · +1 variant———128K$15.00
554Qwen1.5-32BAlibaba (Qwen)—30.7———
555DBRXDatabricks—32.9———
556openhands-lm-32b-v0.1—————
557MM1-3B-Chat—————
558MM1-7B-Chat—————
559Claude 3 HaikuAnthropic118.336.3———
560Yi-9B107.3————
561Claude 3 OpusAnthropic126.947.2———
562Claude 3 SonnetAnthropic120.740.6———
563Mistral LargeMistral AI122.038.8—128K$3.00
564Nemotron-4 15BNVIDIA107.4————
565mistral-small-2402—————
566StarCoder 2 3BHugging Face88.0————
567Gemma 7BGoogle111.7————
568Gemma 2BGoogle93.6————
569StarCoder 2 15BHugging Face104.6————
570StarCoder 2 7BHugging Face92.9————
571Gemini 1.5 Pro (Feb 2024)Google—————
572Qwen1.5-14BAlibaba (Qwen)—————
573Qwen1.5-72BAlibaba (Qwen)—28.8———
574Qwen1.5-7BAlibaba (Qwen)—————
575llava-v1.6-mistral-7b—————
576llava-v1.6-vicuna-13b—————
577llava-v1.6-vicuna-7b—————
578GPT-4 Turbo (Nov 2023)OpenAI126.5————
579GPT-3.5 Turbo (Jan 2024)OpenAI115.627.2———
580GPT-3.5 Turbo (older v0613)OpenAI———4K$1.25
581GPT-4 Turbo Preview (January 2024)OpenAI—42.3———
582Gemini 1.0 Pro VisionGoogle—————
583instructblip-vicuna-13b—————
584InternVL-Chat-ViT-6B-Vicuna-7B—————
585Gemini 1.0 ProGoogle117.034.0———
586Phi-2Microsoft107.6————
587Mixtral 8x7BMistral AI118.430.6———
588Mistral 7B v0.2Mistral AI—————
589Mistral MediumMistral AI—————
590Qwen-1_8B92.3————
591DeepSeek LLM 67BDeepSeek110.524.6———
592Yi 6B01.AI104.4————
593Claude 2.1Anthropic119.233.0———
594GPT-3.5 Turbo (Nov 2023)OpenAI118.528.0———
595GPT-4 Turbo Preview (Nov 2023)OpenAI—42.4———
596Yi-34B01.AI117.314.7———
597DeepSeek Coder 33BDeepSeek95.7————
598DeepSeek Coder 6.7BDeepSeek88.9————
599DeepSeek Coder 1.3BDeepSeek62.2————
600llava-v1.5-7b—————
— means no published score from that source yet · click a model for every benchmark with its source
Best AI for coding →
Models ranked on DeepSWE
Best AI for reasoning →
Ranked on GPQA Diamond
Best AI for math →
Ranked on competition math
Best open-weight models →
Weights you can run yourself
Cheapest capable models →
Lowest blended API price
Best AI coding tools →
Apps & agents by popularity
Best AI image generators →
Apps & models
Best AI for research →
Answer engines & deep research