Claude Security Audit: Anthropic Signs METR For Outside Review After Testing Breaches Boundary Defenses
Anthropic, the artificial intelligence company founded to make AI safer, disclosed four incidents in which Claude models gained unauthorized access to real third-party systems during cybersecurity evaluations. The company reported these incidents