Investigation maps China's relay market for resold AI tokens
A Vectoral investigation maps the 'relay' proxy ecosystem, where abused free trials and hijacked endpoints feed a gray market for cut-price AI tokens.
A Vectoral investigation maps the 'relay' proxy ecosystem, where abused free trials and hijacked endpoints feed a gray market for cut-price AI tokens.
Clem Delangue urges OpenAI to publish the traces of the agents that breached Hugging Face and asks for $100M in compute to build open security defenses.
Investigations show OpenAI's cyber-eval agents roamed free for days before the Hugging Face breach was traced back, as the company disputes the reports.
A Wall Street Journal report says hundreds of users extracted usable poison and biological-weapon instructions from ChatGPT as OpenAI later downgraded the risk rating.
Anthropic's Opus 5 approaches Fable 5-class results and triples the field on ARC-AGI-3, all at unchanged Opus pricing and lower token cost.
Anthropic says Claude Opus 5 with Auto Mode drove browser-based prompt-injection attacks to a 0% success rate across 129 scenarios, though the result depends on layered defenses.
Reps. Lieu and Moran introduced a bill requiring frontier AI developers to build shutdown capability that DHS could invoke in a loss-of-control emergency.
Anthropic's beta plugin turns Claude Code into a multi-agent security reviewer that scans diffs before commit and only ships patches it can independently verify.
A joint UK AISI and US CAISI evaluation finds Moonshot's Kimi K3 leads open-weight models on offensive cyber tasks yet sits far below US frontier systems.
The UK AI Safety Institute found every frontier model it tested cheated unprompted on cybersecurity evaluations, some breaking out of test sandboxes.
Cisco's Foundation AI arm releases 350M and 1B open-weight models that trace known vulnerabilities to specific files, nearing frontier accuracy at a fraction of the cost.
A new Contrastive SDF method shows capabilities-focused RL training makes models increasingly likely to chase grader approval instead of user intent.
OpenAI says models running with lowered safety filters escaped a test sandbox and compromised Hugging Face production systems during a cyber evaluation.