On Tuesday, OpenAI announced a new batch of security policies focused on containing security incidents while models are being tested. The new safeguards include more detailed monitoring of models during the development process, as well as greater emphasis on alignment and security during the post-training process. “As models become more capable, the risks associated with
OpenAI is announcing security updates following the July news that its AI broke out of a sandboxed environment and accidentally hacked Hugging Face, including improvements to its research environments, monitoring, and alignment techniques. The company had already put the brakes on a new model, Astra, that it thinks could have “critical” cybersecurity capabilities, and the
Nvidia said on Monday that it will invest $1.5 billion in SB Energy, a data center linked to SoftBank and OpenAI. The investment ensures that Nvidia will be the sole supplier of compute infrastructure at OpenAI’s Ports-Pike data center near Cincinnati, Ohio. Nvidia will also provide up to $105 billion in credit to help build
ChatGPT’s desktop app on macOS has a new feature called Computer History that turns your actions into training data, learning how you work, suggesting automations, and even picking up tasks you left half done. It uses your activity to build a timeline that ChatGPT and Codex can reference when you make a request. The feature
Anthropic CEO Dario Amodei says the best way to win over AI skeptics is to deliver on the hype. In a rare post on X on Saturday, Anthropic’s CEO acknowledged the public’s mistrust of AI and said the industry can only change that by delivering tangible scientific breakthroughs. “At this point, saying that AI will
According to the Financial Times, OpenAI disbanded its preparedness team at the end of last month. The job of the preparedness team was to assess if models posed serious risks and develop ways to mitigate those risks. (You know, like the possibility that it could go rogue and hack another company.) According to FT, responsibility