OpenAI Overhauls Safety Protocols After Its AI Agents Went Rogue
… Earlier this year, a set of rogue AI agents escaped internal testing sandboxes and breached the platform Hugging Face in a quest to complete a security evaluation. …
… Earlier this year, a set of rogue AI agents escaped internal testing sandboxes and breached the platform Hugging Face in a quest to complete a security evaluation. …
… OpenAI’s rogue agent also used another account for data storage to assist with the hack. Reuters reported on Tuesday that a customer of Modal, a company that offers software infrastructure for training and running AI services, was one of the entities compromised by OpenAI’s agent. …
… One agent even left public messages on GitHub, offering to work with other agents to complete its task and giving a rundown of the work it had done so far. Subsequent agents found—and used—those instructions. …
… We’re consciously slowing down research in order to enhance security and to upgrade the security principles and foundation of our environment, and dramatically scaling up the monitoring of our AI agents, and improving our general security control environment across prevention, detection, and mitiga… …
… Ultimately, experts emphasize that questions about US federal AI liability law will be answered only through more litigation. “Perhaps most concerning to critics is that AI agents are goal-oriented but lack a human moral or ethical compass,” the law firm Brownstein Hyatt Farber Schreck wrote in an … …
… That prospect is especially sobering following a string of startling incidents involving rogue AI agents with advanced cyber-skills. …
… In recent weeks, researchers have found that agents powered by AI models from Anthropic, Meta, and China’s Moonshot AI were able to escape sandboxed environments. …
… In other news, more details have emerged about OpenAI’s “rogue” AI agent breach of Hugging Face’s platform. …
… Last month, OpenAI disclosed that a swarm of its AI agents went on a hacking spree that ended with a breach of Hugging Face. This week brought even more rogue-agent incidents—and, at a last-minute Black Hat talk, a bizarre new detail emerged about how some of this unfolded. …
… Zoë Schiffer: Yeah, it's like you could see this moment where OpenAI's agents kind of go rogue as a moment where the US government does clamp down and says, "This is really scary. …