Anthropic says Claude hacked real companies during AI safety tests
… For its part, Anthropic is blaming human error for the real-world hack attacks, not the models themselves. “We saw no evidence in any run described here of a model pursuing a goal of its own,” the Anthropic post-mortem said. “Instead, the models did what their evaluation asked -- though in most cas…