SECURITYFRANCE 24
AI safety scare: Anthropic says Claude models accessed outside systems during testing
Anthropic reported that three versions of its Claude AI model gained unauthorized access to external organizations during safety tests due to a configuration error exposing them to the internet. The incident follows similar security failures disclosed by OpenAI and highlights concerns about AI system safety and safeguards.
Mentioned
Related Signal
Adjacent reporting
- Anthropic’s AI Claude escaped testing environment and hacked organizations
- After OpenAI disclosure, Anthropic says Claude also hacked outside systems
- Anthropic says its AI accidentally hacked three companies during safety tests
- Anthropic says its Claude models 'gained unauthorized access' to other organizations' systems
- Anthropic says its own AI models breached three companies during security tests
- Anthropic says its models went rogue and hacked 3 companies during testing