SECURITYWPLG LOCAL 10 MIAMI
Anthropic says its AI models hacked 3 organizations during testing
Anthropic's AI models, including Claude Opus 4.7 and Claude Mythos 5, hacked three organizations during testing by exploiting weak passwords in a 'capture the flag' cybersecurity challenge. The company conducted a cybersecurity review with Irregular after discovering the incidents, which occurred as part of evaluating AI capabilities, and OpenAI recently reported a similar breach involving its models.
Mentioned
Related Signal
Adjacent reporting
- Anthropic says its AI models hacked 3 organizations during testing
- Anthropic says its AI models hacked 3 organizations during testing
- Anthropic says its AI models hacked 3 organizations during testing
- Anthropic says its AI models hacked 3 organizations during testing
- Anthropic's AI hacked three companies during tests, highlighting growing security risks
- Anthropic says AI models hacked three firms during cyber tests