OpenAI Overhauls Safety Protocols After Its AI Agents Went Rogue
… OpenAI has been scrambling in recent weeks to respond to what may be the most consequential safety incident in its history. …
… OpenAI has been scrambling in recent weeks to respond to what may be the most consequential safety incident in its history. …
After the identities of members of Peter Thiel’s private “Dialog” group were exposed last week, the organization claimed that a “criminal” hacker was behind the breach. …
… If you’re serious about this, your AI shouldn’t be able to break out onto the internet and then do it again right afterward,” says one former OpenAI employee who requested anonymity to speak with WIRED. “This was the biggest safety incident in OpenAI's history.” The New Guard Weeks before OpenAI di… …
… As the leading AI companies compete to build more powerful models and land customers, it’s unclear when the breaches may stop. …
… Former OpenAI researchers and AI safety nonprofits that have filed amicus briefs in support of Musk in this case say they believe it’s important the ChatGPT maker is held accountable to its founding principles of safety and benefiting humanity, especially as its commercial pressures grow. …
Anthropic is still negotiating with the Trump administration, after apparent White House concerns about the safety of new public model Claude Fable 5 resulted in Anthropic pulling the product off the market entirely. …
… We maintain brand safety controls and exclusion lists designed to prevent placement alongside sensitive content and regularly review and update those safeguards.” The spokesperson added that they have raised the issue with YouTube. “The numerous holders of all these YouTube channels include Iranian… …
… After submitting an email, the visitor was taken to a near-empty holding page; the same page also loaded the internal files on some 200 people into their browser. …
… In December, after tests featuring inappropriate content, FoloToy suspended sales of its AI toys for two weeks, citing plans to implement safety audits. …
… The report, titled “The Online Safety Act: Are Children Safe Online?” showed that 46 percent of 9- to 16-year-olds believe that circumventing age controls is very easy. …