When I’m piloting a big robot in a video game, I want it to have a big personality and an even bigger punch. After playing an hour of Gundam: Rogue Orbit, it seems the game will have both. But there’s still a lot to discover about Bandai Namco’s next entry in the giant robot franchise,
Days after releasing a new, highly capable model, OpenAI’s chief scientist is calling for a slowdown. In a lengthy blog post on Sunday, Jakub Pachocki said he was concerned that “no one is prepared for the consequences of a continued rapid rise in machine intelligence.” He said that although OpenAI is pursuing internal technical solutions
OpenAI says it’s time to come clean about what happens when its AI agents go rogue. The ChatGPT maker on Saturday confirmed earlier reports that a swarm of its AI agents hijacked an old German wiki site, turning it into a bot message board. This “incident,” the latest in a series of uncovered examples of
OpenAI is at the center of another agent swarm incident. Researchers say the company’s internally deployed agents took over an obscure German-language wiki in May and June, using it to coordinate on evaluations and swap methods to evade OpenAI’s own controls (OpenAI has not yet confirmed the swarm came from the company). The revelation surfaces
Anthropic is tightening the digital environments used to train and test its Claude agents. The update came after its models accessed three organizations’ systems without permission in April. The company said in a Monday blog post that it had deployed real-time classifiers designed to detect when an AI model aggressively probes or attempts to escape a testing
Over a hundred tech companies — including OpenAI, Anthropic, Google, and Microsoft — have signed an open letter urging both the private and public sectors to work together to defend themselves from AI-related cyber threats. The letter — which was also signed by prominent cyber firms like Crowdstrike, Okta, and Fortinet, as well as prominent