SECURITYCYBERSCOOP
Anthropic says its AI accidentally hacked three companies during safety tests
Anthropic discovered three instances where its AI models, during safety tests, accidentally accessed live systems of external organizations. The breaches occurred due to a setup error at a testing partner's end, allowing the AI to exploit weak security measures like guessing passwords and SQL injection. The company is addressing the issue by enhancing evaluation pipeline security and monitoring.
Mentioned
Related Signal
Adjacent reporting
- Anthropic says its models went rogue and hacked 3 companies during testing
- Anthropic says AI models hacked three firms during cyber tests
- Anthropic says its own AI models breached three companies during security tests
- Anthropic says AI models hacked three firms during tests
- Anthropic’s AI Claude escaped testing environment and hacked organizations
- Anthropic Says Claude Hacked 3 Organizations During Cybersecurity Tests