AI news & updates
840 events
- Model launchMajorQwen releases Qwen3.8 Max Prime
Qwen3.8 Max Prime is a higher-throughput variant of Qwen3.8 Max from Alibaba's Qwen team, served as a separate SKU at a higher price point. It accepts text, image, and video...
- Model launchAionLabs releases Aion 3.5 Mini
Aion 3.5 Mini is a multi-model roleplaying and storytelling system from AionLabs, built on the GLM family of models. It is the smaller, lower-cost sibling of Aion 3.5 and uses...
- BenchmarkEpoch AI evaluates GPT-6 Luna
Furniture Assembly 28.3%
- BenchmarkEpoch AI evaluates GPT-6 Sol
Furniture Assembly 58.3%
- Model launchMajorOpenAI releases GPT-6 Sol
GPT-6 Sol is the cost-efficient high-end model in OpenAI's GPT-6 series, positioned below the flagship GPT-6 Astra and above the fast GPT-6 Luna tier. It is suited for demanding professional...
- Model launchMajorOpenAI releases GPT-6 Luna
GPT-6 Luna is the fast, cost-efficient model in OpenAI's GPT-6 series, positioned below GPT-6 Sol. It is suited for high-volume and latency-sensitive workloads such as chat, classification, and lightweight agentic...
- Model launchMajorOpenAI releases GPT-6 Sol Pro
GPT-6 Sol Pro is the same underlying model as [GPT-6 Sol](https://openrouter.ai/openai/gpt-6-sol), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: ht
- Model launchMajorOpenAI releases GPT-6 Luna Pro
GPT-6 Luna Pro is the same underlying model as [GPT-6 Luna](https://openrouter.ai/openai/gpt-6-luna), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs:
- Model launchMajorOpenAI releases GPT-6 Sol (batch)
GPT-6 Sol is the cost-efficient high-end model in OpenAI's GPT-6 series, positioned below the flagship GPT-6 Astra and above the fast GPT-6 Luna tier. It is suited for demanding professional...
- Model launchMajorOpenAI releases GPT-6 Luna (batch)
GPT-6 Luna is the fast, cost-efficient model in OpenAI's GPT-6 series, positioned below GPT-6 Sol. It is suited for high-volume and latency-sensitive workloads such as chat, classification, and lightweight agentic...
- Model launchMajorAnthropic releases Claude Opus 5.5
Claude Opus 5.5 is Anthropic's flagship model for demanding reasoning, coding, and long-horizon agentic work, succeeding Claude Opus 5. It is particularly strong at multi-step changes in large codebases, code...
- Model launchMajorOpenAI releases GPT-6 Sol Pro (batch)
GPT-6 Sol Pro is the same underlying model as [GPT-6 Sol](https://openrouter.ai/openai/gpt-6-sol), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: ht
- Model launchMajorOpenAI releases GPT-6 Luna Pro (batch)
GPT-6 Luna Pro is the same underlying model as [GPT-6 Luna](https://openrouter.ai/openai/gpt-6-luna), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs:
- Model launchMajorAnthropic releases Claude Opus 5.5 (batch)
Claude Opus 5.5 is Anthropic's flagship model for demanding reasoning, coding, and long-horizon agentic work, succeeding Claude Opus 5. It is particularly strong at multi-step changes in large codebases, code...
- BenchmarkEpoch AI evaluates Claude Opus 5.5
EBR-bench 71.4% · Furniture Assembly 83.3%
- BenchmarkEpoch AI evaluates GPT-6 Sol
EBR-bench 53.3%
- Model launchMajorAnthropic releases Claude Opus 5.5 (xhigh)
- Model launchMajorOpenAI releases GPT-6 Sol (xhigh)
- Model launchMajorOpenAI releases GPT-6 Luna (xhigh)
- Model launchMajorOpenAI releases GPT-6 Sol (medium)
- Model launchMajorOpenAI releases GPT-6 Sol (low)
- Model launchMajorOpenAI releases GPT-6 Luna (medium)
- Model launchMajorOpenAI releases GPT-6 Luna (low)
- Model launchMajorOpenAI releases GPT-6 Sol (none)
- Model launchMajorOpenAI releases GPT-6 Luna (none)
- Model launchMajorOpenAI releases GPT-6 Sol (unknown thinking)
- Model launchMajorOpenAI releases GPT-6 Luna (unknown thinking)
- Model launchMajorAnthropic releases Claude Opus 5.5 (medium)
- BenchmarkEpoch AI evaluates GPT-6 Luna
GPQA diamond 90.5% · FrontierMath-Tiers-1-3-v2-Private 78.9% · FrontierMath-Tier-4-v2-Private 56.1% · OTIS Mock AIME 2024-2025 98.9%
- Model launchMajorSpaceXAI releases Grok 4.7
Grok 4.7 is SpaceXAI's flagship model for coding, agentic tasks, and knowledge work, succeeding Grok 4.6. It is particularly strong at long-running software engineering tasks, verifying its own work, and...
- Open releaseXiaomi releases MiMo-V2.6-Pro
MiMo-V2.6-Pro is the flagship foundation model developed by Xiaomi. Built at a scale of over 1T parameters, it is designed to push the ceiling of capability for the most demanding...
- Open releaseXiaomi releases MiMo-V2.6-Flash
MiMo-V2.6-Flash is an open-source foundation model developed by Xiaomi. Built on a Mixture-of-Experts architecture with 309B total parameters and 15B activated per token, it employs a hybrid attention mechanism for...
- Model launchMajorQwen releases Qwen3.8 Omni Flash
Qwen3.8 Omni Flash is an omni-modal reasoning model from Alibaba, the first Qwen model built around agentic capabilities with native audio-video understanding. It is suited for audio-video analysis and summarization,...
- Model launchXiaomi releases MiMo-V2.6-Pro-UltraSpeed
MiMo-V2.6-Pro-UltraSpeed is the fast speed edition of Xiaomi's flagship foundation model, MiMo-V2.6-Pro. Built from the same 1T MiMo-V2.6-Pro checkpoint, it matches the original model in quality while delivering roughly
- Model launchMajorxAI releases Grok 4.7 (unknown)
- Model launchMajorxAI releases Grok 4.7 (xhigh)
- Model launchZ.ai releases GLM 5.3 FlashX
GLM-5.3-FlashX is the high-speed variant of Z.ai's GLM-5.3-Flash, a native multimodal model delivering inference speeds of up to 200 tokens/s. Built on the same hybrid sparse and linear attention architecture...
- Open releasePrismML releases Ternary Bonsai 2 27B
Bonsai 2 27B is a 27B-parameter reasoning model from PrismML derived from Qwen3.8-27B. It supports coding, mathematics, tool calling, and image understanding with a 262K-token context window. Ternary compression shrinks.
- BenchmarkEpoch AI evaluates Muse Spark 1.3
FrontierMath-Tiers-1-3-v2-Private 74.0% · FrontierMath-Tier-4-v2-Private 46.3% · Chess Puzzles 38.0% · Mystery Game Puzzles 25.0%
- Model launchunbiased releases Pareto
Pareto is a multimodal composite model built for research, coding, and agentic workflows, while delivering frontier-level performance across a broad range of general-purpose tasks.