AI news & updates
852 events
- Model launchMajorOpenAI releases GPT-5.4 Nano (batch)
GPT-5.4 nano is the most lightweight and cost-efficient variant of the GPT-5.4 family, optimized for speed-critical and high-volume tasks. It supports text and image inputs and is designed for low-latency...
- Model launchMajorOpenAI releases GPT-5.4 Mini (batch)
GPT-5.4 mini brings the core capabilities of GPT-5.4 to a faster, more efficient model optimized for high-throughput workloads. It supports text and image inputs with strong performance across reasoning, coding,...
- Model launchMajorOpenAI releases GPT-5.4 mini (xhigh)
- Model launchMajorOpenAI releases GPT-5.4 mini (none)
- Model launchMajorOpenAI releases GPT-5.4 nano (no thinking)
- Model launchMajorOpenAI releases GPT-5.4 nano (low)
- Model launchMajorOpenAI releases GPT-5.4 mini (low)
- Model launchMajorOpenAI releases GPT-5.4 mini (medium)
- Model launchMajorOpenAI releases GPT-5.4 nano (medium)
- Model launchMajorOpenAI releases GPT-5.4 nano (xhigh)
- Model launchMajorOpenAI releases GPT-5.4 mini (unknown thinking)
- Open releaseMajorMistral releases Mistral Small 4
Mistral Small 4 is the next major release in the Mistral Small family, unifying the capabilities of several flagship Mistral models into a single system. It combines strong reasoning from...
- Open releaseMajorMistral releases Mistral Small 4 (batch)
Mistral Small 4 is the next major release in the Mistral Small family, unifying the capabilities of several flagship Mistral models into a single system. It combines strong reasoning from...
- Model launchZ.ai releases GLM 5 Turbo
GLM-5 Turbo is a new model from Z.ai designed for fast inference and strong performance in agent-driven environments such as OpenClaw scenarios. It is deeply optimized for real-world agent workflows...
- Open releaseMajorNVIDIA releases Nemotron 3 Super
NVIDIA Nemotron 3 Super is a 120B-parameter open hybrid MoE model, activating just 12B parameters for maximum compute efficiency and accuracy in complex multi-agent applications. Built on a hybrid Mamba-Transformer...
- Model launchByteDance Seed releases Seed-2.0-Lite
Seed-2.0-Lite is a versatile, cost‑efficient enterprise workhorse that delivers strong multimodal and agent capabilities while offering noticeably lower latency, making it a practical default choice for most production w
- Model launchMajorOpenAI releases GPT-5.4
GPT-5.4 is OpenAI’s latest frontier model, unifying the Codex and GPT lines into a single system. It features a 1M+ token context window (922K input, 128K output) with support for...
- Model launchMajorOpenAI releases GPT-5.4 Pro
GPT-5.4 Pro is OpenAI's most advanced model, building on GPT-5.4's unified architecture with enhanced reasoning capabilities for complex, high-stakes tasks. It features a 1M+ token context window (922K input, 128K...
- Model launchMajorOpenAI releases GPT-5.4 (batch)
GPT-5.4 is OpenAI’s latest frontier model, unifying the Codex and GPT lines into a single system. It features a 1M+ token context window (922K input, 128K output) with support for...
- Model launchMajorOpenAI releases GPT-5.4 Pro (batch)
GPT-5.4 Pro is OpenAI's most advanced model, building on GPT-5.4's unified architecture with enhanced reasoning capabilities for complex, high-stakes tasks. It features a 1M+ token context window (922K input, 128K...
- Model launchMajorOpenAI releases GPT-5.4 (medium)
- Model launchMajorOpenAI releases GPT-5.4 (none)
- Model launchMajorOpenAI releases GPT-5.4 (low)
- Model launchMajorOpenAI releases GPT-5.4 Pro (xhigh)
- Model launchMajorOpenAI releases GPT-5.4 (xhigh)
- Model launchMajorOpenAI releases GPT-5.4 Pro (web)
- Model launchMajorOpenAI releases GPT-5.4 (unknown thinking)
- Model launchMajorOpenAI releases GPT-5.4 Pro (no thinking)
- Model launchInception releases Mercury 2
Mercury 2 is an extremely fast reasoning LLM, and the first reasoning diffusion LLM (dLLM). Instead of generating tokens sequentially, Mercury 2 produces and refines multiple tokens in parallel, achieving...
- Model launchMajorGoogle releases Gemini 3.1 Flash Lite
Gemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads. It supports text, image, video, audio, and PDF inputs, and is designed for lightweight agentic...
- Model launchMajorGoogle releases Gemini 3.1 Flash Lite Preview
Gemini 3.1 Flash Lite Preview is Google's high-efficiency model optimized for high-volume use cases. It outperforms Gemini 2.5 Flash Lite on overall quality and approaches Gemini 2.5 Flash performance across...
- Model launchByteDance Seed releases Seed-2.0-Mini
Seed-2.0-mini targets latency-sensitive, high-concurrency, and cost-sensitive scenarios, emphasizing fast response and flexible inference deployment. It delivers performance comparable to ByteDance-Seed-1.6, supports 256
- Model launchMajorGoogle releases Nano Banana 2 (Gemini 3.1 Flash Image Preview)
Gemini 3.1 Flash Image Preview, a.k.a. "Nano Banana 2," is Google’s latest state of the art image generation and editing model, delivering Pro-level visual quality at Flash speed. It combines...
- Open releaseMajorQwen releases Qwen3.5-27B
The Qwen3.5 27B native vision-language Dense model incorporates a linear attention mechanism, delivering fast response times while balancing inference speed and performance. Its overall capabilities are comparable to tho
- Open releaseMajorQwen releases Qwen3.5-122B-A10B
The Qwen3.5 122B-A10B native vision-language model is built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency. In terms of
- Model launchMajorQwen releases Qwen3.5-Flash
The Qwen3.5 native vision-language Flash models are built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency. Compared to t
- Model launchMajorGoogle releases Gemini 3.1 Pro Preview Custom Tools
Gemini 3.1 Pro Preview Custom Tools is a variant of Gemini 3.1 Pro that improves tool selection behavior by preventing overuse of a general bash tool when more efficient third-party...
- Model launchMajorAlibaba releases Qwen 3.5 Flash (hosted 35B-A3B)
- Open releaseMajorQwen releases Qwen3.5-9B
Qwen3.5-9B is a multimodal foundation model from the Qwen3.5 family, designed to deliver strong reasoning, coding, and visual understanding in an efficient 9B-parameter architecture. It uses a unified vision-language des
- Open releaseMajorQwen releases Qwen3.5-35B-A3B
The Qwen3.5 Series 35B-A3B is a native vision-language model designed with a hybrid architecture that integrates linear attention mechanisms and a sparse mixture-of-experts model, achieving higher inference efficiency. I