Anthropic Says It’s Taking Claude Fable 5 Offline to Comply With US Government Order
… In the blog post, the company argued that it has implemented strong safeguards to reduce the likelihood of Claude Fable 5’s misuse. …
… In the blog post, the company argued that it has implemented strong safeguards to reduce the likelihood of Claude Fable 5’s misuse. …
… It also committed to taking a more comprehensive approach to its security testing through improved defense-in-depth measures and more carefully designed tests. “Evaluation environments increasingly need to be held to the same security standard as any other system our models run in,” the blog post r…
… Some of the safeguards Anthropic decided on were unsurprising: The company said it would reroute users who asked questions about cybersecurity, biology, or chemistry to a less capable AI model to reduce the chances of someone using the advanced AI to carry out a cyberattack or build a bioweapon. …
… In April, OpenAI also privately launched a model that it said has advanced cybersecurity capabilities and convened a working group similar to Project Glasswing. …