AI news & updates
1,090 events
- BenchmarkEpoch AI evaluates Qwen3.7 Flash
Mystery Game Puzzles 14.0%
- BenchmarkEpoch AI evaluates GPT-5 mini (medium)
Mystery Game Puzzles 4.0%
- BenchmarkEpoch AI evaluates GPT-5 Nano
Mystery Game Puzzles 8.0%
- BenchmarkEpoch AI evaluates GPT-5.4 nano (medium)
Mystery Game Puzzles 6.0%
- BenchmarkEpoch AI evaluates GPT-5.6 Sol (none)
Mystery Game Puzzles 33.0%
- BenchmarkEpoch AI evaluates GPT-5.5 (no thinking)
Mystery Game Puzzles 18.0%
- BenchmarkEpoch AI evaluates gpt-oss-120b
Mystery Game Puzzles 0.0% · Mystery Game Puzzles 2.0%
- BenchmarkEpoch AI evaluates GPT-5.2 (none)
Mystery Game Puzzles 14.0%
- BenchmarkEpoch AI evaluates DeepSeek v4 Pro (unknown thinking)
Mystery Game Puzzles 17.0%
- BenchmarkEpoch AI evaluates GPT-5.4 nano (low)
Mystery Game Puzzles 3.0%
- BenchmarkEpoch AI evaluates Nemotron 3 Ultra
Mystery Game Puzzles 20.0%
- BenchmarkEpoch AI evaluates GPT-5.6 Terra (none)
Mystery Game Puzzles 14.0%
- BenchmarkEpoch AI evaluates GPT-5 nano (medium)
Mystery Game Puzzles 9.0%
- BenchmarkEpoch AI evaluates GPT-5.1 (no thinking)
Mystery Game Puzzles 15.0%
- BenchmarkEpoch AI evaluates GPT-5.4 mini (none)
Mystery Game Puzzles 11.0%
- BenchmarkEpoch AI evaluates GPT-5.4 nano (no thinking)
Mystery Game Puzzles 9.0%
- AnnouncementIntelligent transcription with Gemini 3.5 Transcribe
Now you can get more intelligent speech-to-text transcription with Gemini 3.5 Transcribe.
- Open releaseMajorQwen releases Qwen3.8 Flash
Qwen3.8 Flash is a multimodal reasoning model from Alibaba. It is suited for coding assistance, agentic workflows, visual understanding, document and codebase analysis, desktop interaction, chart analysis, and long-video
- Open releaseZ.ai releases GLM 5.3 Flash (batch)
GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while...
- BenchmarkEpoch AI evaluates GLM 5.3 Flash
GPQA diamond 90.2% · OTIS Mock AIME 2024-2025 93.9% · Chess Puzzles 14.0%
- AnnouncementTraining and Finetuning Multi-Vector Embedding Models with Sentence TransformersHugging Face✓ primaryHugging Face
- Feature5 ways to upgrade your home decor with Google Search
Learn how to use Google Search tools to find home decor inspiration, shop for furniture, and tackle DIY projects.
Google (The Keyword)✓ primaryGoogle - AnnouncementGranite 4.2 LLMs: How They're Built
- AnnouncementQuantization-Aware Healing: a compressed, 4-bit model that outperforms its full-precision originalHugging Face✓ primaryHugging Face
- BenchmarkEpoch AI evaluates GLM 5.3
FrontierMath-Tiers-1-3-v2-Private 68.8% · FrontierMath-Tier-4-v2-Private 29.3%
- AnnouncementWire It, Run It, Deploy It: AI Workflows in GradioHugging Face✓ primaryHugging Face
- AnnouncementMistral x HUMAINMistral AI✓ primaryMistral AI
- BenchmarkEpoch AI evaluates GLM 5.3
GPQA diamond 90.9% · OTIS Mock AIME 2024-2025 91.1% · Chess Puzzles 21.0%
- ResearchFrom Atari to EVE Online: Building on 15 Years of AI Research in Games
Google DeepMind partners with game studios to prototype breakthrough AI gameplay.
Google DeepMind✓ primaryGoogle - Model launchMajorMeta releases Muse Spark 1.2 Contributor
Muse Spark 1.2 contributor tier is a reasoning model from Meta designed for developers who want to start building at an even lower cost. It’s meaningfully cheaper than Muse Spark...
- Open releaseMajorDeepSeek releases DeepSeek V4 Flash Vision Exp
DeepSeek V4 Flash Vision Exp is an experimental vision-enabled version of [DeepSeek V4 Flash 0731](https://openrouter.ai/deepseek/deepseek-v4-flash-0731) from DeepSeek, adding image understanding while matching the base
- AnnouncementHow Hugging Face Inference Endpoints, Jobs, and Buckets Power Search on Papers with CodeHugging Face✓ primaryHugging Face
- BenchmarkMeasuring benchmark optimization in speech recognitionHugging Face✓ primaryHugging Face
- AnnouncementUp to 3.2x Faster Inference with LFM2.5-DSparkHugging Face✓ primaryHugging Face
- AnnouncementAgentic Search. More accurate and efficient results from your AI systems.
The retrieval layer that helps AI systems navigate, read, and verify information inside even the most complex documents
Mistral AI✓ primaryMistral AI - Open releaseZ.ai releases GLM 5.3 Flash
GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while...
- Open releaseTencent releases Hy-MT2-1.8B
Hy-MT2-1.8B is a compact 1.8B-parameter translation model from Tencent. It supports 33 language pairs and five Chinese dialect and minority-language pairs, with workflows for structured, delimiter-based, contextual, glos
- Open releaseTencent releases Hy-MT2-30B-A3B
Hy-MT2-30B-A3B is Tencent's flagship translation model in the Hy-MT2 family. It supports 33 language pairs and five Chinese dialect and minority-language pairs, with workflows for structured, delimiter-based, contextual,
- Open releaseZ.ai (Zhipu AI) releases GLM-5.3-Flash (unknown)
- Research5 new ways to level up your learning with Search
Here’s how you can use Google Search tools to study for classes and standardized tests.
Google (The Keyword)✓ primaryGoogle