AI news & updates
152 events
- Open releaseMajorBaseten on Hugging Face Inference Providers 🔥Hugging Face✓ primaryHugging Face
- Open releaseMajorDeepSeek releases DeepSeek V4 Flash 0731
DeepSeek V4 Flash 0731 is a sparse mixture-of-experts model from DeepSeek, with 13B active parameters out of 284B total. This re-post-trained revision is suited for coding, reasoning, and agent workflows....
- Open releaseMajorDeepSeek releases DeepSeek V4 Flash 0731 (low)
- Open releaseMajorDeepSeek releases DeepSeek V4 Flash 0731 (none)
- Open releaseMajorDeepSeek releases DeepSeek V4 Flash 0731 (unknown)
- Open releaseinclusionAI releases Ling 3.0 Flash
*Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (MoE) model*, with approximately *5.1B parameters activated per token*. The model is designed with *token efficiency and production-scale agentic inference* as key
- Open releasePoolside releases Laguna S 2.1
Laguna S 2.1 is the latest coding agent model from [Poolside](<https://poolside.ai/>). Laguna S 2.1 is a 118B total parameter model with 8B active parameters, scoring 70.2% on Terminal-Bench 2.1 and...
- Open releaseMeituan releases LongCat 2.0
LongCat 2.0 is a sparse mixture-of-experts language model from Meituan, with 48B active parameters out of 1.6T total. It is suited for coding, repository-level changes, long-horizon problem solving, and agentic...
- Open releaseMoonshotAI releases Kimi K3
Kimi K3 is a 2.8T parameter open-weight multimodal reasoning model from Moonshot AI. It is suited for complex coding, knowledge work, and long-horizon agentic workflows, and is particularly strong at...
- Open releaseMoonshotAI releases Kimi K3 (batch)
Kimi K3 is a 2.8T parameter open-weight multimodal reasoning model from Moonshot AI. It is suited for complex coding, knowledge work, and long-horizon agentic workflows, and is particularly strong at...
- Open releaseMoonshot releases Kimi K3 (low)
- Open releaseMoonshot releases Kimi K3 (unknown)
- Open releaseThinking Machines releases Inkling
Inkling is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 41B active parameters out of 975B total. It is designed for general-purpose reasoning, coding, agentic and tool-use systems,.
- Open releaseThinking Machines releases Inkling Small
Inkling Small is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 12B active parameters out of 276B total. It is positioned as the smaller, more efficient member of...
- Open releaseThinking Machines releases Inkling Small (xhigh)
- Open releaseThinking Machines releases Inkling (xhigh)
- Open releaseThinking Machines releases Inkling Small (medium)
- Open releaseThinking Machines releases Inkling Small (low)
- Open releaseThinking Machines releases Inkling Small (minimal)
- Open releaseThinking Machines releases Inkling Small (none)
- Open releaseMajorRun AI workloads on any cloud, store on Hugging Face: zero-egress storage with SkyPilotHugging Face✓ primaryHugging Face
- Open releaseTencent releases Hy3
Hy3 is a 295B-parameter Mixture-of-Experts model from Tencent (21B active, 192 experts with top-8 routing) built for reasoning, agentic workflows, and real-world production use. It supports a configurable reasoning effor
- Open releasePoolside releases Laguna XS 2.1
Laguna XS 2.1 is the latest coding agent model in the 33B-A3B category from [Poolside](https://poolside.ai/) and a step forward from their Laguna XS.2 model (released in April 2026). It combines...
- Open releaseMajorFeaturing Every Eval Ever Results on Hugging Face Model PagesHugging Face✓ primaryHugging Face
- Open releaseZ.ai releases GLM 5.2
GLM 5.2 is a large-scale reasoning model from Z.ai. It supports text input and output with a 1M-token context window, and is suited for long-horizon agent workflows, project-level software engineering,...
- Open releaseMoonshotAI releases Kimi K2.7 Code
MoonshotAI: Kimi K2.7 Code is a coding-focused model in Moonshot AI's Kimi K2 family, built to complete end-to-end programming tasks reliably over long contexts. It uses a native multimodal mixture-of-experts...
- Open releaseMajorNVIDIA releases Nemotron 3 Ultra
NVIDIA Nemotron 3 Ultra is an open frontier-reasoning and orchestration model from NVIDIA, with 55B active parameters out of 550B total (MoE). Built on a hybrid Transformer-Mamba mixture-of-experts architecture, it...
- Open releaseMajorNVIDIA releases Nemotron 3.5 Content Safety
NVIDIA Nemotron 3.5 Content Safety is a compact 4B-parameter multimodal guardrail model from NVIDIA, fine-tuned from Google Gemma-3-4B. It moderates both inputs to and responses from LLMs and VLMs, accepting...
- Open releaseMiniMax releases MiniMax M3
MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding,...
- Open releaseStepFun releases Step 3.7 Flash
Step 3.7 Flash is StepFun's latest high-efficiency multimodal Mixture-of-Experts model. It pairs a 196B-parameter language backbone with a vision encoder for native image and video understanding, activating roughly 11B p
- Open releaseMajorMistral releases Mistral Medium 3.5
Mistral Medium 3.5 is a dense 128B instruction-following model from Mistral AI. It supports text and image inputs with text output, and is designed for agentic workflows, coding, and complex...
- Open releasePoolside releases Laguna XS.2
- Open releaseMajorQwen releases Qwen3.6 35B A3B
Qwen3.6-35B-A3B is an open-weight multimodal model from Alibaba Cloud with 35 billion total parameters and 3 billion active parameters per token. It uses a hybrid sparse mixture-of-experts architecture combining Gated...
- Open releaseMajorDeepSeek releases DeepSeek V4 Pro 0423
DeepSeek V4 Pro is a large-scale Mixture-of-Experts model from DeepSeek with 1.6T total parameters and 49B activated parameters, supporting a 1M-token context window. It is designed for advanced reasoning, coding,...
- Open releaseMajorDeepSeek releases DeepSeek V4 Flash 0423
DeepSeek V4 Flash is an efficiency-optimized Mixture-of-Experts model from DeepSeek with 284B total parameters and 13B activated parameters, supporting a 1M-token context window. It is designed for fast inference and...
- Open releaseMajorDeepSeek releases DeepSeek v4 Pro (unknown thinking)
- Open releaseMajorDeepSeek releases DeepSeek v4 Pro (high)
- Open releaseMajorDeepSeek releases DeepSeek v4 (unknown)
- Open releaseMajorDeepSeek releases DeepSeek v4 Flash (max)
- Open releaseMajorQwen releases Qwen3.6 27B
Qwen3.6 27B is a dense 27-billion-parameter language model from the Qwen Team at Alibaba, released in April 2026. It features hybrid multimodal capabilities — accepting text, image, and video inputs...