This is The Stepback, a weekly newsletter breaking down one essential story from the tech world. For more on AI safety, follow Robert Hart. The Stepback arrives in our subscribers’ inboxes at 8AM ET. Opt in for The Stepback here. It all started in July, when one of OpenAI’s autonomous AI agents went rogue during
Your AI Slop Bores Me is brilliant in its simplicity. There are two tabs: human and LARP as an AI. On one side you enter a request. On the other, you submit an answer. But the important thing is that there’s a human on both sides of the equation. Prompts can request a response as
Anthropic is trying to calm everyone down over its new watermark. The company sparked a wave of discourse this week when it updated a support page for its Claude model to reveal its plans to mark AI-generated text as AI-generated. Now, after concerns from Claude users, Anthropic has published a blog post on Friday that
Anthropic’s move to watermark Claude-generated text is driving some users away from its AI assistant. The AI lab said this week that some versions of Claude will embed an “imperceptible watermark” directly into their text output. It will travel with the content when it is copied and pasted and may survive some editing, Anthropic said,
Google will now allow you to remove visible watermarks from the images, videos, and music made with AI tools. With the update, you can toggle off a new “Media watermark” setting in Gemini and Google’s AI video generator, Flow. When toggled off, Google will remove the “sparkle” watermark that appears in the bottom-right corner of
Another OpenAI executive is departing. Denise Dresser, who joined OpenAI as its chief revenue officer in December after serving as CEO of Slack, will be leaving in the “coming weeks” to “pursue other opportunities,” she said in a team note posted to LinkedIn. Dali Rajic, president and COO of Wiz, will be taking over the