IndustryMistral introduced EU data processing and paid priority access, although regional routing excludes several features and priority availability is not guaranteed.The Decoder ↗
AgentsTencent Hunyuan presented research evaluating whether self-modifying AI agents genuinely improve, targeting more reliable measurement of self-evolution.Tencent Hunyuan on X ↗
AIGCXiaomi's MiLM Plus introduced PROVE, combining two perception-aligned object-removal metrics with a real-world video benchmark.MarkTechPost ↗
AgentsAEROBAT automated agent-behavior research, generating 79 hypotheses and executing 23,512 simulation rounds across 12 target behaviors.arXiv ↗
AgentsMESA improved long-horizon agent-memory performance by 8.5% while using 41% fewer evidence tokens than an all-structure approach.arXiv ↗
ResearchFACT trains world-action models on failed actions, reducing success-biased future hallucinations and improving simulated and real-world bimanual manipulation.arXiv ↗
SafetyMarkNull reduced AI-image watermark bit accuracy to 53.14% without perceptible degradation and also compromised Google's SynthID-Image.arXiv ↗
SafetyA deterministic streaming guardrail withheld every chunk completing predefined dangerous lexical pairs, but researchers stressed its coverage remains deliberately narrow.arXiv ↗
SafetyResearchers catalogued LLM-mediated variants of eight classic web attacks, finding vulnerability depends on both application architecture and model behavior.arXiv ↗
ResearchAn empirical study examined when chain-of-thought improves reasoning and when sequential processing depth instead becomes a performance bottleneck.arXiv ↗
AgentsResearchers proposed closed-loop LLM copilots that combine agricultural sensor data, domain knowledge and feedback-driven recommendations for digital farming.arXiv ↗
AIGCQwen made its Qwen-Image-3.0 image generator available for testing through OpenArt.Alibaba Qwen on X ↗
ModelsQwen said Qwen3.8-Max climbed from 22nd to fourth place on the Legal Research Bench leaderboard.Alibaba Qwen on X ↗
Mistral will host rival open models, beginning with Z.ai’s GLM-5.2, while adding regional inference and reserved European capacity.
⚡ Quick hits
AgentsArtificial Analysis found GPT-5.5 xhigh led its new AnalystAgent benchmark in single-run reliability with a 66% pass rate.Artificial Analysis on X ↗
LTX’s new open-weight world model adds adaptive rendering, native 4K HDR output and stronger multi-shot control for local video production.
⚡ Quick hits
AIGCIdeogram launched Ad Resizer, converting one uploaded design into campaign-ready assets for social, video and connected-TV placements.Ideogram on X ↗
Google’s AI assistant crossed one billion monthly users, matching ChatGPT and becoming its fastest-growing product to reach that scale.
⚡ Quick hits
AIGCRunway added Seedance 2.5, supporting 50 character references and music-synchronized video clips lasting up to 30 seconds.Runway on X ↗
AgentsReplit expanded its MCP beta so external agents can create, load, search and publish Replit applications directly.Replit on X ↗
FundingGeneral Catalyst led a $1.1 billion funding round for River AI, a startup founded only two months earlier.TechCrunch ↗
AgentsOpenAI released a Linux preview of its ChatGPT desktop app, bringing ChatGPT Work and Codex to several major distributions.OpenAI on X ↗
ResearchGoogle demonstrated AMIE conducting real-time audio-visual clinical consultations in a first-of-its-kind research study.Google Research ↗
AIGCKrea launched a Slack beta combining multiple generative models with an agent that adapts to a creative team’s preferences.Krea on X ↗
AIGCLuma introduced Scenes, letting creators refine individual scenes without regenerating an entire AI-produced advertisement.Luma AI on X ↗
IndustryElectricSQL, creator of browser-embeddable Postgres engine PGlite, is joining Databricks to bring WASM Postgres into AI-agent sandboxes.Databricks on X ↗
Open SourceUnsloth released an open-source desktop application for running and training models locally across macOS, Windows and Linux.Unsloth on X ↗
The open 30B mixture-of-experts model activates 3B parameters per token and ships with training data, recipes and a new model router.
⚡ Quick hits
IndustryManus said it will resume operating independently and announced transition measures required by regulations in specified regions.Manus on X ↗
AgentsLlamaIndex introduced ExtractBench, a benchmark for measuring information extraction from complex enterprise documents.LlamaIndex on X ↗
PolicySpotify will label AI Persona artist profiles and exclude their tracks from algorithmic recommendations, while leaving them searchable.TechCrunch ↗