Skip to content

Newest models

Most recent releases, with whatever independent scores exist so far.

Most recent releases, with whatever independent scores exist so far. Source: OpenRouter / Epoch AI. Variants (reasoning effort, batch) are folded into one row.

#ModelCapabilityReasoningCodingContext$/M
451Amazon Nova LiteAmazon—————
452Amazon Nova MicroAmazon—————
453Yi-Lightning01.AI—————
454INTELLECT-1Prime Intellect,Hugging Face,Arcee AI100.4————
455Tulu 3 (Tülu 3) 70BAllen Institute for AI—46.3———
456GPT-4o (Nov 2024)OpenAI128.847.9———
457GPT-4o (2024-11-20)OpenAI———128K$4.38
458Mistral Large 2407Mistral AI———131K$3.00
459Mistral Large 2 (Nov 2024)Mistral AI128.551.3———
460Qwen2.5 Coder 32B InstructAlibaba (Qwen)———33K$0.74
461UnslopNemo 12BTheDrummer———1.02M$0.40
462Qwen2.5-Coder-14B-Instruct—————
463Qwen2.5-Coder-3B-Instruct—————
464Qwen-TurboAlibaba (Qwen)—41.8———
465Claude 3.5 Sonnet (October 2024)Anthropic133.6————
466Claude 3.5 HaikuAnthropic127.2————
467Claude 3.5 Haiku (Oct 2024)Anthropic—38.1———
468Claude 3.5 Sonnet (Oct 2024)Anthropic—55.3———
469Magnum v4 72Banthracite-org———33K$3.13
470Ministral 3BMistral AI118.125.3———
471Ministral 8BMistral AI—27.1———
472Qwen2.5 7B InstructAlibaba (Qwen)———33K$0.13
473Gemini 1.5 Flash 8BGoogle—33.0———
474Llama 3.2 1B InstructMeta———60K$0.070
475Llama 3.2 3B InstructMeta———131K$0.12
476Gemini 1.5 Pro (Sept 2024)Google131.757.2———
477Gemini 1.5 Flash (Sep 2024)Google129.447.3———
478Llama 3.2 90BMeta125.541.0———
479Llama 3.2 1BMeta102.023.9———
480Llama 3.2 11BMeta—————
481Llama 3.2 3BMeta—————
482Qwen2.5-72BAlibaba (Qwen)129.049.1———
483Qwen2.5-7BAlibaba (Qwen)118.435.5———
484Qwen2.5 72B InstructAlibaba (Qwen)———33K$0.37
485Qwen2.5-14BAlibaba (Qwen)—————
486Qwen2.5-Coder-32BAlibaba (Qwen)119.4————
487Qwen2.5-Coder-14B116.3————
488Qwen2.5-Coder (7B)Alibaba (Qwen)112.9————
489Qwen2.5-Coder-3B107.4————
490Qwen2.5-Coder (1.5B)Alibaba (Qwen)102.6————
491Qwen2.5-Coder-0.5B87.4————
492Qwen2.5-32BAlibaba (Qwen)128.546.1———
493Mistral Small v24.09Mistral AI—————
494Pixtral 12BMistral AI—————
495o1-miniOpenAI · +2 variants135.862.4———
496o1-previewOpenAI134.850.3———
497DeepSeek-V2.5 (Sep 2024)DeepSeek—————
498Command R+Cohere119.3————
499c4ai-command-r-08-2024—————
500Command R (08-2024)Cohere———128K$0.26
— means no published score from that source yet · click a model for every benchmark with its source
Best AI for coding →
Models ranked on DeepSWE
Best AI for reasoning →
Ranked on GPQA Diamond
Best AI for math →
Ranked on competition math
Best open-weight models →
Weights you can run yourself
Cheapest capable models →
Lowest blended API price
Best AI coding tools →
Apps & agents by popularity
Best AI image generators →
Apps & models
Best AI for research →
Answer engines & deep research