Skip to content

Newest models

Most recent releases, with whatever independent scores exist so far.

Most recent releases, with whatever independent scores exist so far. Source: OpenRouter / Epoch AI. Variants (reasoning effort, batch) are folded into one row.

#ModelCapabilityReasoningCodingContext$/M
601Qwen-7BAlibaba (Qwen)106.5————
602GPT-3.5 Turbo InstructOpenAI———4K$1.63
603Mistral 7B v0.1Mistral AI112.0————
604Qwen-14BAlibaba (Qwen)112.8————
605Baichuan 2-7BBaichuan95.8————
606internlm-20b111.8————
607internlm-chat-20b—————
608Phi-1.5Microsoft90.8————
609Falcon-180BTechnology Innovation Institute111.9————
610Baichuan2-13BBaichuan102.8————
611Baichuan2-13B-Chat—————
612GPT-3.5 Turbo 16kOpenAI———16K$3.25
613Qwen-VL-Chat—————
614Claude InstantAnthropic120.2————
615Weaver (alpha)Mancer———8K$0.49
616ReMM SLERP 13Bundi95———6K$0.42
617Stable Beluga 2Stability AI117.0————
618Llama 2-70BMeta113.626.3———
619Llama 2-13BMeta105.8————
620Llama 2-34BMeta104.8————
621Llama 2-7BMeta98.5————
622Claude 2Anthropic120.134.7———
623Baichuan 1-13BBaichuan—————
624internlm-7b102.4————
625MythoMax 13Bgryphe———8K$0.087
626XGen-7BSalesforce92.8————
627chatglm2-6b98.4————
628MPT-30BMosaicML100.2————
629Inflection-1Inflection AI—————
630Vicuna-13B-v1.3Large Model Systems Organization,University of California (UC) Berkeley—————
631GPT-4 (Jun 2023)OpenAI123.130.7———
632GPT-3.5 Turbo (Jun 2023)OpenAI113.1————
633GPT-3.5 Turbo (June 2023)OpenAI—————
634open_llama_7b91.0————
635Baichuan1-7BBaichuan89.8————
636GPT-3.5 TurboOpenAI · +1 variant———16K$0.75
637GPT-4OpenAI———8K$37.50
638Falcon-40BTechnology Innovation Institute104.0————
639instructblip-vicuna-7b—————
640PaLM 2-L114.8————
641PaLM 2-M108.0————
642PaLM 2-S105.8————
643MPT-7BMosaicML94.0————
644RedPajama-INCITE-7B-Base89.7————
645Falcon-7BTechnology Innovation Institute94.5————
646stablelm-tuned-alpha-7b54.3————
647Claude 1.3Anthropic—————
648vicuna-13b-v1.194.0————
649Dolly 2.0-12bDatabricks88.9————
650Cerebras-GPT-13BCerebras82.4————
— means no published score from that source yet · click a model for every benchmark with its source
Best AI for coding →
Models ranked on DeepSWE
Best AI for reasoning →
Ranked on GPQA Diamond
Best AI for math →
Ranked on competition math
Best open-weight models →
Weights you can run yourself
Cheapest capable models →
Lowest blended API price
Best AI coding tools →
Apps & agents by popularity
Best AI image generators →
Apps & models
Best AI for research →
Answer engines & deep research