AI news & updates
848 events
- Open releaseMajorQwen releases Qwen3 VL 8B Thinking
Qwen3-VL-8B-Thinking is the reasoning-optimized variant of the Qwen3-VL-8B multimodal model, designed for advanced visual and textual reasoning across complex scenes, documents, and temporal sequences. It integrates enha
- Open releaseMajorQwen releases Qwen3 VL 8B Instruct
Qwen3-VL-8B-Instruct is a multimodal vision-language model from the Qwen3-VL series, built for high-fidelity understanding and reasoning across text, images, and video. It features improved multimodal fusion with Interle
- Model launchMajorOpenAI releases GPT-5 Pro
GPT-5 Pro is OpenAI’s most advanced model, offering major improvements in reasoning, code quality, and user experience. It is optimized for complex tasks that require step-by-step reasoning, instruction following, and...
- Model launchMajorGoogle releases Nano Banana (Gemini 2.5 Flash Image)
Gemini 2.5 Flash Image, a.k.a. "Nano Banana," is now generally available. It is a state of the art image generation model with contextual understanding. It is capable of image generation,...
- Model launchMajorOpenAI releases GPT-5 Pro (batch)
GPT-5 Pro is OpenAI’s most advanced model, offering major improvements in reasoning, code quality, and user experience. It is optimized for complex tasks that require step-by-step reasoning, instruction following, and...
- Open releaseMajorQwen releases Qwen3 VL 30B A3B Thinking
Qwen3-VL-30B-A3B-Thinking is a multimodal model that unifies strong text generation with visual understanding for images and videos. Its Thinking variant enhances reasoning in STEM, math, and complex tasks. It excels...
- Open releaseMajorQwen releases Qwen3 VL 30B A3B Instruct
Qwen3-VL-30B-A3B-Instruct is a multimodal model that unifies strong text generation with visual understanding for images and videos. Its Instruct variant optimizes instruction-following for general multimodal tasks. It e
- Model launchreleases tiny-recursion-modelEpoch AI Benchmarking Hubtiny-recursion-model
- Open releaseZ.ai releases GLM 4.6
Compared with GLM-4.5, this generation brings several key improvements: Longer context window: The context window has been expanded from 128K to 200K tokens, enabling the model to handle more complex...
- Open releaseZ.ai (Zhipu AI),Tsinghua University releases GLM-4.6 (Together)
- Open releaseMajorDeepSeek releases DeepSeek V3.2 Exp
DeepSeek-V3.2-Exp is an experimental large language model released by DeepSeek as an intermediate step between V3.1 and future architectures. It introduces DeepSeek Sparse Attention (DSA), a fine-grained sparse attention
- Model launchMajorAnthropic releases Claude Sonnet 4.5
Claude Sonnet 4.5 is Anthropic’s most advanced Sonnet model to date, optimized for real-world agents and coding workflows. It delivers state-of-the-art performance on coding benchmarks such as SWE-bench Verified, with...
- Model launchMajorAnthropic releases Claude Sonnet 4.5 (batch)
Claude Sonnet 4.5 is Anthropic’s most advanced Sonnet model to date, optimized for real-world agents and coding workflows. It delivers state-of-the-art performance on coding benchmarks such as SWE-bench Verified, with...
- Model launchMajorAnthropic releases Claude Sonnet 4.5 (16k thinking)
- Model launchMajorAnthropic releases Claude Sonnet 4.5 (59k thinking)
- Model launchMajorAnthropic releases Claude Sonnet 4.5 (32k thinking)
- Model launchMajorAnthropic releases Claude Sonnet 4.5 (no thinking)
- Model launchMajorAnthropic releases Claude Sonnet 4.5 (48k thinking)
- Model launchMajorAnthropic releases Claude Sonnet 4.5 (8k thinking)
- Model launchMajorAnthropic releases Claude Sonnet 4.5 (1k thinking)
- Model launchMajorAnthropic releases Claude Sonnet 4.5 (unknown thinking)
- Model launchMajorAnthropic releases Claude Sonnet 4.5 (2k thinking)
- Model launchMajorAnthropic releases Claude Sonnet 4.5 (low)
- Model launchMajorAnthropic releases Claude Sonnet 4.5 (12k thinking)
- Open releaseTheDrummer releases Cydonia 24B V4.1
Uncensored and creative writing model based on Mistral Small 3.2 24B with good recall, prompt adherence, and intelligence.
- Model launchRelace releases Relace Apply 3
Relace Apply 3 is a specialized code-patching LLM that merges AI-suggested edits straight into your source files. It can apply updates from GPT-4o, Claude, and others into your files at...
- Model launchreleases gemini-robotics-er-1.5-previewEpoch AI Benchmarking Hubgemini-robotics-er-1.5-preview
- Model launchMajorGoogle DeepMind releases Gemini 2.5 Flash (Sep 2025)
- Model launchMajorGoogle DeepMind releases Gemini 2.5 Flash-Lite (Sep 2025)
- Model launchMajorQwen releases Qwen3 Max
Qwen3-Max is an updated release built on the Qwen3 series, offering major improvements in reasoning, instruction following, multilingual support, and long-tail knowledge coverage compared to the January 2025 version. It.
- Model launchMajorAlibaba releases Qwen3-Max-Instruct
- Model launchMajorQwen releases Qwen3 Coder Plus
Qwen3 Coder Plus is Alibaba's proprietary version of the Open Source Qwen3 Coder 480B A35B. It is a powerful coding agent model specializing in autonomous programming via tool calling and...
- Open releaseMajorQwen releases Qwen3 VL 235B A22B Thinking
Qwen3-VL-235B-A22B Thinking is a multimodal model that unifies strong text generation with visual understanding across images and video. The Thinking model is optimized for multimodal reasoning in STEM and math....
- Open releaseMajorQwen releases Qwen3 VL 235B A22B Instruct
Qwen3-VL-235B-A22B Instruct is an open-weight multimodal model that unifies strong text generation with visual understanding across images and video. The Instruct model targets general vision-language use (VQA, document
- Open releaseMajorDeepSeek releases DeepSeek V3.1 Terminus
DeepSeek-V3.1 Terminus is an update to [DeepSeek V3.1](/deepseek/deepseek-chat-v3.1) that maintains the model's original capabilities while addressing issues reported by users, including language consistency and agent ca
- Model launchMajorxAI releases Grok 4 Fast
- Open releaseMajorMistral AI releases Magistral Small 1.2
- Model launchMajorQwen releases Qwen3 Coder Flash
Qwen3 Coder Flash is Alibaba's fast and cost efficient version of their proprietary Qwen3 Coder Plus. It is a powerful coding agent model specializing in autonomous programming via tool calling...
- Model launchMajorOpenAI releases GPT-5-codex
- Open releaseMajorQwen releases Qwen3 Next 80B A3B Thinking
Qwen3-Next-80B-A3B-Thinking is a reasoning-first chat model in the Qwen3-Next line that outputs structured “thinking” traces by default. It’s designed for hard multi-step problems; math proofs, code synthesis/debugging,