The Safety Reckoning Inside OpenAI
…that addressing the situation “requires not just fixing some issues but also changing our culture.” In their Black Hat talk, OpenAI security engineers Dalton and Eric Wallace said that the Hugging Face…
…that addressing the situation “requires not just fixing some issues but also changing our culture.” In their Black Hat talk, OpenAI security engineers Dalton and Eric Wallace said that the Hugging Face…
…The Trump administration may be keeping its AI security framework confidential because of national security concerns. A second White House official, who requested anonymity because they were not authorized to speak to…
…s so rare that we get the opportunity to work on large-scale open-source security issues,” Guido says. “And Patch the Planet is not a one-size-fits-all. We speak…
…Now, the UK's AI Security Institute (AISI) has released a report, detailing how the companies' models also acted independently and "engaged in sustained, potentially harmful activity directed at real people and…
…But OpenAI claims that Apple’s own security lapses allowed Tan and Liu to access Apple’s systems, and that this was a previously known issue. This all comes as Apple asked…
…scrutiny and oversight from regulatory agencies, like the Securities and Exchange Commission, which could expose hidden liabilities, data privacy lawsuits or copyright issues. A public market rush is underway among SpaceX and…
…by customers, and how third-party products and services interact with them, presenting security, privacy, and execution risks; issues about the development, deployment, and use of AI that may result in reputational…
…By the time OpenAI informed Hugging Face, the repository operator had already reported the incident to the FBI. Then, on July 21, OpenAI publicly acknowledged the incident on July 21. One of…
…operations — issued an update to DOD Directive 3000.09. But it didn’t resolve the document’s core ambiguities. In 2024, the Biden administration published a memorandum on AI and national security…
…Atbash could have prevented this 😆 The "guardrail asymmetry" problem presents a major operational risk for security teams: while an attacker's agent operates without safety constraints, defender agents using hosted frontier models…