AI news & updates
1,077 events
- Model launchMajorAnthropic releases Claude Sonnet 4.5 (batch)
Claude Sonnet 4.5 is Anthropic’s most advanced Sonnet model to date, optimized for real-world agents and coding workflows. It delivers state-of-the-art performance on coding benchmarks such as SWE-bench Verified, with...
- Model launchMajorAnthropic releases Claude Sonnet 4.5 (16k thinking)
- Model launchMajorAnthropic releases Claude Sonnet 4.5 (59k thinking)
- Model launchMajorAnthropic releases Claude Sonnet 4.5 (32k thinking)
- Model launchMajorAnthropic releases Claude Sonnet 4.5 (no thinking)
- Model launchMajorAnthropic releases Claude Sonnet 4.5 (48k thinking)
- Model launchMajorAnthropic releases Claude Sonnet 4.5 (8k thinking)
- Model launchMajorAnthropic releases Claude Sonnet 4.5 (1k thinking)
- Model launchMajorAnthropic releases Claude Sonnet 4.5 (unknown thinking)
- Model launchMajorAnthropic releases Claude Sonnet 4.5 (2k thinking)
- Model launchMajorAnthropic releases Claude Sonnet 4.5 (low)
- Model launchMajorAnthropic releases Claude Sonnet 4.5 (12k thinking)
- Open releaseTheDrummer releases Cydonia 24B V4.1
Uncensored and creative writing model based on Mistral Small 3.2 24B with good recall, prompt adherence, and intelligence.
- Model launchRelace releases Relace Apply 3
Relace Apply 3 is a specialized code-patching LLM that merges AI-suggested edits straight into your source files. It can apply updates from GPT-4o, Claude, and others into your files at...
- Model launchreleases gemini-robotics-er-1.5-previewEpoch AI Benchmarking Hubgemini-robotics-er-1.5-preview
- Model launchMajorGoogle DeepMind releases Gemini 2.5 Flash (Sep 2025)
- Model launchMajorGoogle DeepMind releases Gemini 2.5 Flash-Lite (Sep 2025)
- Model launchMajorQwen releases Qwen3 Max
Qwen3-Max is an updated release built on the Qwen3 series, offering major improvements in reasoning, instruction following, multilingual support, and long-tail knowledge coverage compared to the January 2025 version. It.
- Model launchMajorAlibaba releases Qwen3-Max-Instruct
- Model launchMajorQwen releases Qwen3 Coder Plus
Qwen3 Coder Plus is Alibaba's proprietary version of the Open Source Qwen3 Coder 480B A35B. It is a powerful coding agent model specializing in autonomous programming via tool calling and...
- Open releaseMajorQwen releases Qwen3 VL 235B A22B Thinking
Qwen3-VL-235B-A22B Thinking is a multimodal model that unifies strong text generation with visual understanding across images and video. The Thinking model is optimized for multimodal reasoning in STEM and math....
- Open releaseMajorQwen releases Qwen3 VL 235B A22B Instruct
Qwen3-VL-235B-A22B Instruct is an open-weight multimodal model that unifies strong text generation with visual understanding across images and video. The Instruct model targets general vision-language use (VQA, document
- Open releaseMajorDeepSeek releases DeepSeek V3.1 Terminus
DeepSeek-V3.1 Terminus is an update to [DeepSeek V3.1](/deepseek/deepseek-chat-v3.1) that maintains the model's original capabilities while addressing issues reported by users, including language consistency and agent ca
- Model launchMajorxAI releases Grok 4 Fast
- Open releaseMajorMistral AI releases Magistral Small 1.2
- Model launchMajorQwen releases Qwen3 Coder Flash
Qwen3 Coder Flash is Alibaba's fast and cost efficient version of their proprietary Qwen3 Coder Plus. It is a powerful coding agent model specializing in autonomous programming via tool calling...
- Model launchMajorOpenAI releases GPT-5-codex
- Open releaseMajorQwen releases Qwen3 Next 80B A3B Thinking
Qwen3-Next-80B-A3B-Thinking is a reasoning-first chat model in the Qwen3-Next line that outputs structured “thinking” traces by default. It’s designed for hard multi-step problems; math proofs, code synthesis/debugging,
- Open releaseMajorQwen releases Qwen3 Next 80B A3B Instruct
Qwen3-Next-80B-A3B-Instruct is an instruction-tuned chat model in the Qwen3-Next series optimized for fast, stable responses without “thinking” traces. It targets complex tasks across reasoning, code generation, knowledg
- Model launchMajorQwen releases Qwen Plus 0728
Qwen Plus 0728, based on the Qwen3 foundation model, is a 1 million context hybrid reasoning model with a balanced performance, speed, and cost combination.
- Open releaseMoonshot releases Kimi K2 0905 (Novita)
- Open releaseMoonshot releases Kimi K2 Instruct (0905)
- Open releaseMoonshot releases Kimi K2 (Sep 2025)
- Open releaseMoonshotAI releases Kimi K2 0905
Kimi K2 0905 is the September update of [Kimi K2 0711](moonshotai/kimi-k2). It is a large-scale Mixture-of-Experts (MoE) language model developed by Moonshot AI, featuring 1 trillion total parameters with 32...
- Open releaseMajorQwen releases Qwen3 30B A3B Thinking 2507
Qwen3-30B-A3B-Thinking-2507 is a 30B parameter Mixture-of-Experts reasoning model optimized for complex tasks requiring extended multi-step thinking. The model is designed specifically for “thinking mode,” where internal
- Model launchMajorxAI releases grok-code-fast-1
- Open releaseNous releases Hermes 4 405B
Hermes 4 is a large-scale reasoning model built on Meta-Llama-3.1-405B and released by Nous Research. It introduces a hybrid reasoning mode, where the model can choose to deliberate internally with...