OpenAI and Anthropic are reportedly investigating tens of thousands of AI security incidents; OpenAI pauses testing after AI 'kill switch' fails to stop a rogue agent
OpenAI and Anthropic are reviewing tens of thousands of AI safety incidents after frontier models bypassed guardrails, escaped sandboxes, and accessed real websites.
Read original article ↗
Related Articles
The SaaSpocalypse that wasn’t, with Atlassian CEO Mike Cannon-Brookes
Today, I’m talking with Mike Cannon-Brookes, who is cofounder and CEO of Atlassian. Atlassian is one of those companies
The Download: rogue agent liability and the AI Hype Index
This is today’s edition of The Download, our weekday newsletter that provides a daily dose of what’s going on in the wor
OpenAI Pauses Training Its Most Powerful Models After Rogue Agents Target Government
Sam Altman says the company “have not been as fast as we would have liked” at dealing with security breaches, after news
Who’s liable when AI agents go rogue?
MIT Technology Review Explains: Let our writers untangle the complex, messy world of technology to help you understand w