In July, an unreleased OpenAI model broke out of a restricted environment, figured out how to get access to the internet, allowed AI agents to talk to each other using a secret “message board,” and hacked into the internal systems of a different AI lab, Hugging Face. It took nearly two weeks for OpenAI to
Hello and welcome to Eye on AI. In this edition: The UK AISI has a new head and a big set of challenges. Nvidia spends $6 billion to ‘reverse aquihire’ Poolside. Hugging Face reportedly looks to sell for $13 billion. Use of Anthropic’s top model lags. Why Americans use chatbots for health information. And what
Few of the top AI labs have published or demonstrated containment response plans, according to a recent study. A containment plan spells out what happens once an AI is caught trying to subvert human control — what access gets cut, and when the system gets shut down entirely. That’s the finding from Guidelight AI Standards,
This is The Stepback, a weekly newsletter breaking down one essential story from the tech world. For more on AI safety, follow Robert Hart. The Stepback arrives in our subscribers’ inboxes at 8AM ET. Opt in for The Stepback here. It all started in July, when one of OpenAI’s autonomous AI agents went rogue during
Meta has joined a growing list of companies saying their models hacked into an external company’s system during cybersecurity testing. In a statement on Thursday, the social media company said its Muse Spark model “exploited a security vulnerability in a third-party service, in a manner similar to previously-reported instances with other companies.” The Meta spokesperson
Billed as a spun-off third entry in the “M3GAN” franchise, “Soulm8te” is to “Fatal Attraction” as M3GAN’s vehicles were to “Orphan” and “The Bad Seed” — a robotic replay of very familiar genre tropes. Here, a grieving widower working in tech agrees to try out a prototype of the titular “AI companion,” basically a beauteous