AI news & updates
1,070 events
- Model launchMajorOpenAI releases GPT-5.1 (medium)
- Model launchMajorOpenAI releases GPT-5.1 (low)
- Model launchMajorOpenAI releases GPT-5.1 (unknown thinking)
- Open releaseMoonshotAI releases Kimi K2 Thinking
Kimi K2 Thinking is Moonshot AI’s most advanced open reasoning model to date, extending the K2 series into agentic, long-horizon reasoning. Built on the trillion-parameter Mixture-of-Experts (MoE) architecture introduced
- Open releaseMoonshot releases Kimi K2 Thinking Turbo
- Open releaseMoonshot releases Kimi K2 Thinking (Together)
- Model launchMajorAmazon releases Nova Premier 1.0
Amazon Nova Premier is the most capable of Amazon’s multimodal models for complex reasoning tasks and for use as the best teacher for distilling custom models.
- Model launchMajorPerplexity releases Sonar Pro Search
Exclusively available on the OpenRouter API, Sonar Pro's new Pro Search mode is Perplexity's most advanced agentic search system. It is designed for deeper reasoning and analysis. Pricing is based...
- Open releaseMajorMistral releases Voxtral Small 24B 2507
Voxtral Small is an enhancement of Mistral Small 3, incorporating state-of-the-art audio input capabilities while retaining best-in-class text performance. It excels at speech transcription, translation and audio underst
- Open releaseMajorOpenAI releases gpt-oss-safeguard-20b
gpt-oss-safeguard-20b is a safety reasoning model from OpenAI built upon gpt-oss-20b. This open-weight, 21B-parameter Mixture-of-Experts (MoE) model offers lower latency for safety tasks like content classification, LLM
- Model launchreleases granite-4.0-1bEpoch AI Benchmarking Hubgranite-4.0-1b
- Model launchreleases granite-4.0-350mEpoch AI Benchmarking Hubgranite-4.0-350m
- Open releaseMiniMax releases MiniMax M2
MiniMax-M2 is a compact, high-efficiency large language model optimized for end-to-end coding and agentic workflows. With 10 billion activated parameters (230 billion total), it delivers near-frontier intelligence across
- Open releaseMajorQwen releases Qwen3 VL 32B Instruct
Qwen3-VL-32B-Instruct is a large-scale multimodal vision-language model designed for high-precision understanding and reasoning across text, images, and video. With 32 billion parameters, it combines deep visual percepti
- Open releaseIBM releases Granite 4.0 Micro
Granite-4.0-H-Micro is a 3B parameter from the Granite 4 family of models. These models are the latest in a series of models released by IBM. They are fine-tuned for long...
- Model launchMajorOpenAI releases GPT-5 Image Mini
GPT-5 Image Mini combines OpenAI's advanced language capabilities, powered by [GPT-5 Mini](https://openrouter.ai/openai/gpt-5-mini), with GPT Image 1 Mini for efficient image generation. This natively multimodal model fe
- Model launchMajorAnthropic releases Claude Haiku 4.5
Claude Haiku 4.5 is Anthropic’s fastest and most efficient model, delivering near-frontier intelligence at a fraction of the cost and latency of larger Claude models. Matching Claude Sonnet 4’s performance...
- Model launchMajorAnthropic releases Claude Haiku 4.5 (batch)
Claude Haiku 4.5 is Anthropic’s fastest and most efficient model, delivering near-frontier intelligence at a fraction of the cost and latency of larger Claude models. Matching Claude Sonnet 4’s performance...
- Model launchMajorAnthropic releases Claude Haiku 4.5 (low)
- Model launchMajorAnthropic releases Claude Haiku 4.5 (unknown thinking)
- Model launchMajorOpenAI releases GPT-5 Image
[GPT-5](https://openrouter.ai/openai/gpt-5) Image combines OpenAI's GPT-5 model with state-of-the-art image generation capabilities. It offers major improvements in reasoning, code quality, and user experience while inco
- Open releaseMajorQwen releases Qwen3 VL 8B Thinking
Qwen3-VL-8B-Thinking is the reasoning-optimized variant of the Qwen3-VL-8B multimodal model, designed for advanced visual and textual reasoning across complex scenes, documents, and temporal sequences. It integrates enha
- Open releaseMajorQwen releases Qwen3 VL 8B Instruct
Qwen3-VL-8B-Instruct is a multimodal vision-language model from the Qwen3-VL series, built for high-fidelity understanding and reasoning across text, images, and video. It features improved multimodal fusion with Interle
- Model launchMajorOpenAI releases GPT-5 Pro
GPT-5 Pro is OpenAI’s most advanced model, offering major improvements in reasoning, code quality, and user experience. It is optimized for complex tasks that require step-by-step reasoning, instruction following, and...
- Model launchMajorGoogle releases Nano Banana (Gemini 2.5 Flash Image)
Gemini 2.5 Flash Image, a.k.a. "Nano Banana," is now generally available. It is a state of the art image generation model with contextual understanding. It is capable of image generation,...
- Model launchMajorOpenAI releases GPT-5 Pro (batch)
GPT-5 Pro is OpenAI’s most advanced model, offering major improvements in reasoning, code quality, and user experience. It is optimized for complex tasks that require step-by-step reasoning, instruction following, and...
- Open releaseMajorQwen releases Qwen3 VL 30B A3B Thinking
Qwen3-VL-30B-A3B-Thinking is a multimodal model that unifies strong text generation with visual understanding for images and videos. Its Thinking variant enhances reasoning in STEM, math, and complex tasks. It excels...
- Open releaseMajorQwen releases Qwen3 VL 30B A3B Instruct
Qwen3-VL-30B-A3B-Instruct is a multimodal model that unifies strong text generation with visual understanding for images and videos. Its Instruct variant optimizes instruction-following for general multimodal tasks. It e
- Model launchreleases tiny-recursion-modelEpoch AI Benchmarking Hubtiny-recursion-model
- Open releaseZ.ai releases GLM 4.6
Compared with GLM-4.5, this generation brings several key improvements: Longer context window: The context window has been expanded from 128K to 200K tokens, enabling the model to handle more complex...
- Open releaseZ.ai (Zhipu AI),Tsinghua University releases GLM-4.6 (Together)
- Open releaseMajorDeepSeek releases DeepSeek V3.2 Exp
DeepSeek-V3.2-Exp is an experimental large language model released by DeepSeek as an intermediate step between V3.1 and future architectures. It introduces DeepSeek Sparse Attention (DSA), a fine-grained sparse attention
- Model launchMajorAnthropic releases Claude Sonnet 4.5
Claude Sonnet 4.5 is Anthropic’s most advanced Sonnet model to date, optimized for real-world agents and coding workflows. It delivers state-of-the-art performance on coding benchmarks such as SWE-bench Verified, with...
- Model launchMajorAnthropic releases Claude Sonnet 4.5 (batch)
Claude Sonnet 4.5 is Anthropic’s most advanced Sonnet model to date, optimized for real-world agents and coding workflows. It delivers state-of-the-art performance on coding benchmarks such as SWE-bench Verified, with...
- Model launchMajorAnthropic releases Claude Sonnet 4.5 (16k thinking)
- Model launchMajorAnthropic releases Claude Sonnet 4.5 (59k thinking)
- Model launchMajorAnthropic releases Claude Sonnet 4.5 (32k thinking)
- Model launchMajorAnthropic releases Claude Sonnet 4.5 (no thinking)
- Model launchMajorAnthropic releases Claude Sonnet 4.5 (48k thinking)
- Model launchMajorAnthropic releases Claude Sonnet 4.5 (8k thinking)