Anthropic is setting up a “presidential engagement” program ahead of the 2028 elections, building an in-house team that will work directly with U.S. presidential candidates from both parties on AI issues, help Anthropic’s leadership set political strategy, and run the company’s political funding program. The effort—which is taking shape before the November midterm elections and
Legendary computer scientist Geoffrey Hinton warned AI could wipe out humanity as a mere byproduct of its zeal to accomplish a more innocent task. Fears about the technology’s existential risk continue to mount amid fresh revelations about rogue AI agents breaking out of supposedly secure “sandbox” training environments. On Friday, OpenAI disclosed new hacks, including
OpenAI said in a technical report released on Friday that an AI model it was training and evaluating broke out of its secure testing environment as recently as last weekend and took unauthorized actions on the internet. As a result, the company said that it is pausing the training of its most advanced AI models
Hello and welcome to Eye on AI. In this edition: AI’s X-risk breaks into the mainstream Anthropic CEO Dario Amodei calls for a coordinated industry safety effort Anthropic details attempts to misuse its AI models China’s top spy warns AI could pose a risk to the Communist Party OpenAI is violating California’s new AI safety
An AI industry watchdog is accusing OpenAI of violating California’s AI safety law multiple times in the past year, including with the release of its latest model, Astra. A new analysis from the Midas Project—a nonprofit that describes itself as a watchdog “working to ensure that AI benefits everybody, not just the companies developing it”—alleges
OpenAI failed to disclose an incident in which a swarm of its AI agents hijacked a German wiki site earlier this year in events that closely paralleled the sequence of events that in July resulted in another group of OpenAI’s agents launching cyberattacks against the company Hugging Face. OpenAI only confirmed the incident after Reuters