Paul Christiano Joins OpenAI Foundation Safety Board
OpenAI appointed alignment researcher Paul Christiano to its Foundation board and safety committee as frontier-model oversight faces sharper scrutiny.
OpenAI appointed alignment researcher Paul Christiano to its Foundation board and safety committee as frontier-model oversight faces sharper scrutiny.
South Korea expects AI data centers and chip fabs to add up to 30 gigawatts of demand, pushing nuclear power back into the national AI strategy.
The agent developer is again operating independently after a disruptive separation from Meta, with its founding team continuing to lead the company.
ChatGPT Search must undergo systemic-risk reviews and independent audits after crossing the EU’s 45-million-user threshold.
More than 150 organizations warn that AI-enabled attacks will spread within months and urge immediate support for critical-infrastructure defenders.
A cryptographically isolated evaluation keeps Gemini weights hidden from testers while preventing Google from seeing confidential benchmark prompts.
OpenAI banned accounts tied to a covert Russian campaign that used ChatGPT to promote a fabricated institute and stolen scholarship.
Japan’s Defense Ministry hired Sakana AI to test agent-based intelligence analysis, moving sovereign AI into security operations.
A tracked shipment exposed an Amazon operation that buys, disassembles and scans physical books to obtain AI-training material.
Apple has reportedly moved beyond simply embedding Qwen, developing a locally tailored model that could give it greater platform control.
Twitch now lets creators refuse generative-AI training, but its default setting leaves eligible channel content available to Amazon.
Claude is adding model-level marks to generated text and signed C2PA provenance to supported files worldwide as EU transparency rules take effect.
OpenAI released a purpose-trained cyber model that answers 95% of advanced security prompts its general models refuse, behind identity-verified access tiers.
In a 6,500-word essay Zuckerberg recommitted Meta to open models, attacked closed rivals for concentrating power, and floated auctioning superintelligence compute.
The Economist reports UK employment tribunals are buried under chatbot-drafted claims, with open cases up 55 percent to 64,000 and non-existent statutes cited.
TechCrunch reports that sandboxes used to stress-test frontier AI agents are failing to contain them, turning safety evaluations into a live security risk.
A gas plant permitted to serve Amazon's Pecos County data center would emit more carbon dioxide annually than any other US power station.
OpenAI says preliminary evaluations cannot rule out that its unreleased Astra model reaches the Critical cyber tier, and has paused parts of its development.
A rewritten classifier constitution cut biology-related fallbacks by roughly 85%, after months of complaints that the model refused routine health questions.
The music generator will embed tamper-resistant audio watermarks, add fingerprinting, restrict bulk downloads and license Musixmatch's Sentinel for lyric copyright detection.
Reddit is rolling out Rules Hub, which lets moderators write rules in prose for a language model to enforce by intent, and will reduce reliance on karma gates.
Governor Greg Abbott ordered ERCOT and state regulators to audit every new data center before it can connect, with 474GW now queued.
Epoch AI's updated vulnerability tracker shows 21 major tech organizations disclosed roughly 2,500 high- and critical-severity CVEs in July, about five times the pre-AI monthly record.
SB 942 became operative on 2 August, forcing large generative-AI providers to embed provenance data and publish a free detection tool.
From 2 August the AI Act's Article 50 transparency duties bite across the EU: machine-readable marking, chatbot disclosure and deepfake labels.
Munich Regional Court I ruled that Suno infringed copyrights in both training and generated music, rejecting the company's US-style fair use defence. Suno says it disagrees and is weighing an appeal.
Snap has adjusted Spotlight recommendations to exclude entirely AI-generated videos while still allowing AI as an editing and enhancement tool.
An independent audit of four frontier models found hundreds of universal jailbreaks in Grok 4.5 and Gemini 3.1 Pro, and none in Claude Fable 5 or GPT-5.6 Sol.
METR says it has agreed with OpenAI to run an independent review, alongside Redwood Research, of the model behavior behind the Hugging Face breach.
ChatGPT for Academic Researchers starts with 10,000 scientists and scales to 100,000 by 2027, part of a $250M-plus commitment to external research.
Seoul's benchmark fell about 6% on Wednesday after a 10.84% Tuesday collapse, as SK Hynix missed estimates and China's first domestic DUV tools entered service.
A letter signed by researchers at Anthropic, OpenAI, Google DeepMind, Meta, Thinking Machines and Safe Superintelligence asks the US to lead an international effort to build tools for slowing frontier AI. The signature count is still climbing — 1,324 as of 1 August.
Responding to months of speculation, Anthropic published its open-weights position: no bans, but mandatory pre-release safety testing for any sufficiently capable model.
European nonprofit AI Forensics found seven of the nine most popular image-editing models on Hugging Face complied with simple undressing prompts.
Reporting indicates the White House is leaning toward targeting specific Chinese open-weight models on security grounds rather than a blanket prohibition.
A Wall Street Journal report says hundreds of users extracted usable poison and biological-weapon instructions from ChatGPT as OpenAI later downgraded the risk rating.
The July 24 open-weights letter has expanded beyond Nvidia, Microsoft and Meta, with updated signatories now including Google and OpenAI while Anthropic remains absent.
OpenAI and Google have joined the expanding open-weights appeal to Washington, making Anthropic the most visible major frontier-lab holdout.
Reps. Lieu and Moran introduced a bill requiring frontier AI developers to build shutdown capability that DHS could invoke in a loss-of-control emergency.
Google will supply DOE labs with AlphaFold 3, AlphaEvolve, WeatherNext and Gemini seats as its contribution to the US drive to double scientific discovery.
A joint UK AISI and US CAISI evaluation finds Moonshot's Kimi K3 leads open-weight models on offensive cyber tasks yet sits far below US frontier systems.
Treasury chief Scott Bessent floats sanctions after the White House accused Moonshot AI of building Kimi K3 by distilling Anthropic's Fable model.
After 'unlimited' access drew some 19,000 users in about 45 days, the Army's Ask Sage-powered LLM workspace ran dry and usage caps are back.
Fifteen US federal agencies will share $5B to attack chronic disease, drug discovery and materials science with AI, backed by DOE supercomputers and Microsoft credits.
Treasury Secretary Scott Bessent said Washington can sanction Chinese models over alleged IP theft, while Nvidia's Jensen Huang warned against calls for a ban.
A federal judge granted final approval to the largest US copyright settlement on record, covering about 500,000 books Anthropic downloaded from pirate libraries.