Skip to content

Newest models

Most recent releases, with whatever independent scores exist so far.

Most recent releases, with whatever independent scores exist so far. Source: OpenRouter / Epoch AI. Variants (reasoning effort, batch) are folded into one row.

#ModelCapabilityReasoningCodingContext$/M
601DeepSeek Coder 1.3BDeepSeek62.2————
602llava-v1.5-7b—————
603Qwen-7BAlibaba (Qwen)106.5————
604GPT-3.5 Turbo InstructOpenAI———4K$1.63
605Mistral 7B v0.1Mistral AI112.0————
606Qwen-14BAlibaba (Qwen)112.8————
607Baichuan 2-7BBaichuan95.8————
608internlm-20b111.8————
609internlm-chat-20b—————
610Phi-1.5Microsoft90.8————
611Falcon-180BTechnology Innovation Institute111.9————
612Baichuan2-13BBaichuan102.8————
613Baichuan2-13B-Chat—————
614GPT-3.5 Turbo 16kOpenAI———16K$3.25
615Qwen-VL-Chat—————
616Claude InstantAnthropic120.2————
617Weaver (alpha)Mancer———8K$0.49
618ReMM SLERP 13Bundi95———6K$0.42
619Stable Beluga 2Stability AI117.0————
620Llama 2-70BMeta113.626.3———
621Llama 2-13BMeta105.8————
622Llama 2-34BMeta104.8————
623Llama 2-7BMeta98.5————
624Claude 2Anthropic120.134.7———
625Baichuan 1-13BBaichuan—————
626internlm-7b102.4————
627MythoMax 13Bgryphe———8K$0.087
628XGen-7BSalesforce92.8————
629chatglm2-6b98.4————
630MPT-30BMosaicML100.2————
631Inflection-1Inflection AI—————
632Vicuna-13B-v1.3Large Model Systems Organization,University of California (UC) Berkeley—————
633GPT-4 (Jun 2023)OpenAI123.130.7———
634GPT-3.5 Turbo (Jun 2023)OpenAI113.1————
635GPT-3.5 Turbo (June 2023)OpenAI—————
636open_llama_7b91.0————
637Baichuan1-7BBaichuan89.8————
638GPT-3.5 TurboOpenAI · +1 variant———16K$0.75
639GPT-4OpenAI———8K$37.50
640Falcon-40BTechnology Innovation Institute104.0————
641instructblip-vicuna-7b—————
642PaLM 2-L114.8————
643PaLM 2-M108.0————
644PaLM 2-S105.8————
645MPT-7BMosaicML94.0————
646RedPajama-INCITE-7B-Base89.7————
647Falcon-7BTechnology Innovation Institute94.5————
648stablelm-tuned-alpha-7b54.3————
649Claude 1.3Anthropic—————
650vicuna-13b-v1.194.0————
— means no published score from that source yet · click a model for every benchmark with its source
Best AI for coding →
Models ranked on DeepSWE
Best AI for reasoning →
Ranked on GPQA Diamond
Best AI for math →
Ranked on competition math
Best open-weight models →
Weights you can run yourself
Cheapest capable models →
Lowest blended API price
Best AI coding tools →
Apps & agents by popularity
Best AI image generators →
Apps & models
Best AI for research →
Answer engines & deep research