What we learned mapping a year’s worth of AI-enabled cyber threats
… Why it’s harder to assess an actor’s threat level How do security teams assess the risk level of a cyberattacker? …
… Why it’s harder to assess an actor’s threat level How do security teams assess the risk level of a cyberattacker? …
Policy Progress from our Frontier Red Team Mar 19, 2025 In this post, we are sharing what we have learned about the trajectory of potential national security risks from frontier AI models, along with some of our thoughts about challenges and best practices in evaluating these risks. …
… Department of Energy DOE ’s National Nuclear Security Administration NNSA to assess our models for nuclear proliferation risks and continue to work with them on these evaluations. …
… Governments need time to build the capacity to capture benefits while containing risks, which is why they need to start now. Policy that meets the moment We’ve long argued that frontier AI companies should be transparent about what their models can do and how they’re managing the risks. …
… In summary, working with experts , we found that models might soon present risks to national security, if unmitigated. However, we also found that there are mitigations to substantially reduce these risks. We are now scaling up this work in order to reliably identify risks and build mitigations. …
… We both are committed to advancing US national security and defending the American people, and agree on the urgency of applying AI across the government. …
… As AI models become more capable, we believe that they will create major economic and social value, but will also present increasingly severe risks. Our RSP focuses on catastrophic risks – those where an AI model directly causes large scale devastation. …
… For example, techniques that cause Claude to reveal its system prompt are not cybersecurity risks and we do not intend to prevent these types of interactions we even publish them ourselves . …
… Protectionist bans would not address my most serious national security concerns. …