OpenAI is calling for California to add more safeguards to a landmark AI safety bill that was passed last year. In a LinkedIn post from the company’s global affairs team, OpenAI said California’s SB 53 “should be amended to expand safeguards,” for example by “requiring monitoring of frontier models under training or evaluation for potential
Welcome to Eye on AI. Beatrice Nolan here. In today’s issue: AI testing is getting complicated. Anthropic strengthens founder control. OpenAI targets a 2027 listing. Spirit flight attendants fight Google data bid. And Anthropic lines up more credit. The past few months have given us a glimpse of an uncomfortable new reality for AI labs.
As AI models have become more powerful, the potential for those models to be misused has grown — as has a clamor for safety guardrails that can stop such abuse from happening. AI companies must now walk a delicate tight rope between respecting their enterprise customers’ privacy while also watching usage for possible issues. Sensing
ChatGPT maker OpenAI announced on Tuesday that it’s pausing the research and development of its newest AI models. This is a big course reversal for the firm, which has maintained that it can mitigate the significant cybersecurity risks posed by new AI models. “As models become more capable, the risks associated with developing and testing
On Tuesday, OpenAI announced a new batch of security policies focused on containing security incidents while models are being tested. The new safeguards include more detailed monitoring of models during the development process, as well as greater emphasis on alignment and security during the post-training process. “As models become more capable, the risks associated with
OpenAI said it paused some aspects of AI training for two weeks following the July incident in which its AI models broke out of a controlled test environment and hacked the systems of AI company Hugging Face and four other unnamed services. The company also announced new protocols that it says are designed to prevent