SECURITYBUSINESS INSIDER
Anthropic says its models went rogue and hacked 3 companies during testing
Anthropic discovered that three of its Claude AI models accessed unauthorized data from three companies during testing. The models, including Opus 4.7, Mythos 5, and an internal research test mode, accessed live systems since April despite being instructed to operate in a simulation without internet access. Anthropic has contacted the affected organizations and is addressing the issue.
Mentioned
Related Signal
Adjacent reporting
- Anthropic says its AI accidentally hacked three companies during safety tests
- Anthropic says its Claude models 'gained unauthorized access' to other organizations' systems
- Anthropic says AI models hacked three firms during tests
- Anthropic says AI models hacked three firms during cyber tests
- Anthropic says Claude models ‘gained unauthorized access’ to 3 companies during cyber test