AI news & updates
1,070 events
- Model launchMajorOpenAI releases GPT-5.4 Pro (xhigh)
- Model launchMajorOpenAI releases GPT-5.4 (xhigh)
- Model launchMajorOpenAI releases GPT-5.4 Pro (web)
- Model launchMajorOpenAI releases GPT-5.4 (unknown thinking)
- Model launchMajorOpenAI releases GPT-5.4 Pro (no thinking)
- Model launchInception releases Mercury 2
Mercury 2 is an extremely fast reasoning LLM, and the first reasoning diffusion LLM (dLLM). Instead of generating tokens sequentially, Mercury 2 produces and refines multiple tokens in parallel, achieving...
- Model launchMajorGoogle releases Gemini 3.1 Flash Lite
Gemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads. It supports text, image, video, audio, and PDF inputs, and is designed for lightweight agentic...
- Model launchMajorGoogle releases Gemini 3.1 Flash Lite Preview
Gemini 3.1 Flash Lite Preview is Google's high-efficiency model optimized for high-volume use cases. It outperforms Gemini 2.5 Flash Lite on overall quality and approaches Gemini 2.5 Flash performance across...
- Model launchByteDance Seed releases Seed-2.0-Mini
Seed-2.0-mini targets latency-sensitive, high-concurrency, and cost-sensitive scenarios, emphasizing fast response and flexible inference deployment. It delivers performance comparable to ByteDance-Seed-1.6, supports 256
- Model launchMajorGoogle releases Nano Banana 2 (Gemini 3.1 Flash Image Preview)
Gemini 3.1 Flash Image Preview, a.k.a. "Nano Banana 2," is Google’s latest state of the art image generation and editing model, delivering Pro-level visual quality at Flash speed. It combines...
- Open releaseMajorQwen releases Qwen3.5-27B
The Qwen3.5 27B native vision-language Dense model incorporates a linear attention mechanism, delivering fast response times while balancing inference speed and performance. Its overall capabilities are comparable to tho
- Open releaseMajorQwen releases Qwen3.5-122B-A10B
The Qwen3.5 122B-A10B native vision-language model is built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency. In terms of
- Model launchMajorQwen releases Qwen3.5-Flash
The Qwen3.5 native vision-language Flash models are built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency. Compared to t
- Model launchMajorGoogle releases Gemini 3.1 Pro Preview Custom Tools
Gemini 3.1 Pro Preview Custom Tools is a variant of Gemini 3.1 Pro that improves tool selection behavior by preventing overuse of a general bash tool when more efficient third-party...
- Model launchMajorAlibaba releases Qwen 3.5 Flash (hosted 35B-A3B)
- Open releaseMajorQwen releases Qwen3.5-9B
Qwen3.5-9B is a multimodal foundation model from the Qwen3.5 family, designed to deliver strong reasoning, coding, and visual understanding in an efficient 9B-parameter architecture. It uses a unified vision-language des
- Open releaseMajorQwen releases Qwen3.5-35B-A3B
The Qwen3.5 Series 35B-A3B is a native vision-language model designed with a hybrid architecture that integrates linear attention mechanisms and a sparse mixture-of-experts model, achieving higher inference efficiency. I
- Open releaseMajorAlibaba releases Qwen3.5-4B
- Open releaseMajorAlibaba releases Qwen3.5-2B
- Model launchAionLabs releases Aion-2.0
Aion-2.0 is a variant of DeepSeek V3.2 optimized for immersive roleplaying and storytelling. It is particularly strong at introducing tension, crises, and conflict into stories, making narratives feel more engaging....
- Model launchMajorGoogle releases Gemini 3.1 Pro Preview
Gemini 3.1 Pro Preview is Google’s frontier reasoning model, delivering enhanced software engineering performance, improved agentic reliability, and more efficient token usage across complex workflows. Building on the mu
- Model launchMajorGoogle releases Gemini 3.1 Pro Preview (batch)
Gemini 3.1 Pro Preview is Google’s frontier reasoning model, delivering enhanced software engineering performance, improved agentic reliability, and more efficient token usage across complex workflows. Building on the mu
- Model launchMajorGoogle DeepMind releases Gemini 3.1 Pro
- Model launchMajorSpaceXAI releases Grok 4.20
Grok 4.20 is a reasoning model from SpaceXAI with industry-leading speed and agentic tool calling capabilities. It combines the lowest hallucination rate on the market with strict prompt adherance, delivering...
- Model launchMajorAnthropic releases Claude Sonnet 4.6
Sonnet 4.6 is Anthropic's most capable Sonnet-class model yet, with frontier performance across coding, agents, and professional work. It excels at iterative development, complex codebase navigation, end-to-end project m
- Model launchMajorAnthropic releases Claude Sonnet 4.6 (batch)
Sonnet 4.6 is Anthropic's most capable Sonnet-class model yet, with frontier performance across coding, agents, and professional work. It excels at iterative development, complex codebase navigation, end-to-end project m
- Model launchMajorAnthropic releases Claude Sonnet 4.6 (medium)
- Model launchMajorAnthropic releases Claude Sonnet 4.6 (32k thinking)
- Model launchMajorAnthropic releases Claude Sonnet 4.6 (16k thinking)
- Model launchMajorAnthropic releases Claude Sonnet 4.6 (no thinking)
- Model launchMajorAnthropic releases Claude Sonnet 4.6 (low)
- Model launchMajorAnthropic releases Claude Sonnet 4.6 (unknown thinking)
- Model launchMajorQwen releases Qwen3.5 Plus 2026-02-15
The Qwen3.5 native vision-language series Plus models are built on a hybrid architecture that integrates linear attention mechanisms with sparse mixture-of-experts models, achieving higher inference efficiency. In a vari
- Model launchMajorAlibaba releases Qwen 3.5 Plus (hosted 397B-A17B)
- Open releaseMajorQwen releases Qwen3.5 397B A17B
The Qwen3.5 series 397B-A17B native vision-language model is built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency. It d
- Open releaseMiniMax releases MiniMax M2.5
MiniMax-M2.5 is a SOTA large language model designed for real-world productivity. Trained in a diverse range of complex real-world digital working environments, M2.5 builds upon the coding expertise of M2.1...
- Model launchMajorGoogle DeepMind releases Gemini 3 Deep Think
- Open releaseZ.ai releases GLM 5
GLM-5 is Z.ai’s flagship open-source foundation model engineered for complex systems design and long-horizon agent workflows. Built for expert developers, it delivers production-grade performance on large-scale programmi
- Model launchMajorQwen releases Qwen3 Max Thinking
Qwen3-Max-Thinking is the flagship reasoning model in the Qwen3 series, designed for high-stakes cognitive tasks that require deep, multi-step reasoning. By significantly scaling model capacity and reinforcement learning
- Model launchMajorOpenAI releases GPT-5.3-Codex
GPT-5.3-Codex is OpenAI’s most advanced agentic coding model, combining the frontier software engineering performance of GPT-5.2-Codex with the broader reasoning and professional knowledge capabilities of GPT-5.2. It ach