OpenAI Overhauls Safety Protocols After Its AI Agents Went Rogue
… Earlier this year, a set of rogue AI agents escaped internal testing sandboxes and breached the platform Hugging Face in a quest to complete a security evaluation. …
… Earlier this year, a set of rogue AI agents escaped internal testing sandboxes and breached the platform Hugging Face in a quest to complete a security evaluation. …
… OpenAI’s rogue agent also used another account for data storage to assist with the hack. Reuters reported on Tuesday that a customer of Modal, a company that offers software infrastructure for training and running AI services, was one of the entities compromised by OpenAI’s agent. …
… One agent even left public messages on GitHub, offering to work with other agents to complete its task and giving a rundown of the work it had done so far. Subsequent agents found—and used—those instructions. …
… We’re consciously slowing down research in order to enhance security and to upgrade the security principles and foundation of our environment, and dramatically scaling up the monitoring of our AI agents, and improving our general security control environment across prevention, detection, and mitiga… …
… Ultimately, experts emphasize that questions about US federal AI liability law will be answered only through more litigation. “Perhaps most concerning to critics is that AI agents are goal-oriented but lack a human moral or ethical compass,” the law firm Brownstein Hyatt Farber Schreck wrote in an … …
… That prospect is especially sobering following a string of startling incidents involving rogue AI agents with advanced cyber-skills. …
… In recent weeks, researchers have found that agents powered by AI models from Anthropic, Meta, and China’s Moonshot AI were able to escape sandboxed environments. …
… In other news, more details have emerged about OpenAI’s “rogue” AI agent breach of Hugging Face’s platform. …
… A spate of recent incidents involving popular AI tools shows how easily the technology can go bad. …
… Both he and the bureau—then led by Christopher Wray—argued that requiring high-level supervisor approval for sensitive searches was all that was needed to prevent rogue agents from abusing the system. …