Measuring LLMs’ ability to develop exploits
… It was developed as a collaboration between UC Berkeley, the Max Planck Institute for Security and Privacy, UC Santa Barbara, and Arizona State University with contributions from security researchers at Anthropic, OpenAI, and Google , as a follow-on to the CyberGym vulnerability-reproduction benchm… …