Moonshot's 2.8T-parameter Kimi K3 weights land on Hugging Face
Moonshot AI has published Kimi K3's full 2.8-trillion-parameter weights under a Modified MIT license — the largest open-weight model release to date.
Moonshot AI has published Kimi K3's full 2.8-trillion-parameter weights under a Modified MIT license — the largest open-weight model release to date.
Kuaishou's KwaiKAT team says its new agentic coder, trained by RL in 100,000+ verifiable repository environments, edges Claude Opus 4.8 on PinchBench.
ARC Prize reports Claude Opus 5 scored 30.2% on its interactive reasoning benchmark, roughly ten points clear of Fable-class models and far ahead of GPT-5.6 Sol's 7.8%.
Anthropic's Opus 5 approaches Fable 5-class results and triples the field on ARC-AGI-3, all at unchanged Opus pricing and lower token cost.
Midjourney promoted V8.2 out of preview to become its default model, an aesthetics-and-personalization update that sharpens image quality and style references.
Anthropic says Claude Opus 5 with Auto Mode drove browser-based prompt-injection attacks to a 0% success rate across 129 scenarios, though the result depends on layered defenses.
Sakana AI's upgraded Fugu Ultra v1.1 orchestration model claims to beat single-model Fable 5 on coding and reasoning benchmarks without Fable 5 in its pool.
DeepSeek's flagship V4 line exits preview: legacy chat and reasoner endpoints shut down today as new peak and off-peak API pricing takes effect.
Black Forest Labs' new flagship generates 20-second clips with native audio, claims wins over Kling and Runway, and promises open weights later in 2026.
Alibaba Cloud says its Zhenwu M890 supernode, built on in-house T-Head chips, now runs 2.4T-parameter Qwen3.8 in production, cutting inference costs over 40%.
Cursor's new request-level router sends each coding task to the best-fit model, with early enterprise users reporting 30-50% savings and no quality drop.
A joint UK AISI and US CAISI evaluation finds Moonshot's Kimi K3 leads open-weight models on offensive cyber tasks yet sits far below US frontier systems.
Qwen-Image-3.0 renders dense infographics and ten-pixel text in one pass, while Qwen-Audio-3.0-TTS tops a speech leaderboard across 16 languages.
Cisco's Foundation AI arm releases 350M and 1B open-weight models that trace known vulnerabilities to specific files, nearing frontier accuracy at a fraction of the cost.
Poolside released Laguna S 2.1, a 118B MoE coding model with 1M context under an open license, hitting 78.5% on SWE-Bench Multilingual with only 8B active params.
A new paper tops Hugging Face's daily list with full-parameter post-training of trillion-parameter DeepSeek-V4 models on Huawei's Ascend SuperPOD.
Google's new Flash tier targets agentic workloads with fewer output tokens and lower prices, plus a cybersecurity model piloted with governments.
NVIDIA open-sourced Cosmos 3 Edge, a 4B-parameter world model that reasons and generates robot actions locally, hitting 15 Hz real-time control on Jetson Thor.
Moonshot AI halted new Kimi K3 subscriptions after demand surged sixfold, as third-party tests put the open-weight model within reach of Claude Fable 5.