Agentic AI tools now need real safeguards against this kind of indirection.
… This case suggests that agentic AI tools built on large language models may need far stronger runtime safeguards. …
Given the security considerations, it seems unlikely that Claude Mythos will be publicy available anytime soon. However, Anthropic has already stated that it intends Mythos-level features and capabilities to become more widely available. For this to happen, however, it needs to develop and apply various safeguards. The most likely scenario is that public access to Mythos is eventually unblocked, with restrictions to specific features that relate to cybersecurity and biological/medical tasks and research. Project Glasswing is already in operation with a handful of key partners (including Amazon
'You're giving ballistic missiles to individuals with Mythos': JPMorgan CEO Jamie Dimon says Anthropic's AI model poses some serious risks… This case suggests that agentic AI tools built on large language models may need far stronger runtime safeguards. …
… Fable's safeguards are " substantially more effective than those of any previously deployed model" according to Anthropic. Apparently, those safeguards aren't enough for now. …
… For this to happen, however, it needs to develop and apply various safeguards. The most likely scenario is that public access to Mythos is eventually unblocked, with restrictions to specific features that relate to cybersecurity and biological/medical tasks and research. …
… "To address this class of risk, we use a defense-in-depth strategy with safeguards that block malicious instructions at multiple points and help keep tasks aligned with users’ requests. We are continuously strengthening these safeguards as the technology and threat landscape evolve. …
… We also implemented additional safeguards to further enhance the security of information in our possession and to help prevent similar incidents from occurring in the future.” So far, there is no evidence of the data being used in follow-up attacks, or being offered on the dark web. …
BT is the first UK firm to join Anthropic's Project Glasswing High-risk Claude Mythos Preview model is limited to select partners only UK infrastructure set for major security boost – BT already blocks 4m attacks daily BT has become the first UK company to publicly confirm membership of Anthropic’s… …
… Anthropic said Fable 5 had been released with safeguards designed to prevent misuse in cybersecurity attacks, while the full Mythos 5 was kept under tighter controls because of its more advanced capabilities. …
Smart home security cameras are nothing new. …
… New study warns execs are 'knowingly bypassing safeguards because the perceived benefits outweigh the risks'
… By flagging anomalies such as unexpected data transfers or unusual access patterns, UEBA enables security teams to respond swiftly before significant damage occurs. Just as important is having strong technical safeguards in place. …