Skip to content
The Nexus
SECURITYJul 31 · 01:13 UTCCYBERSCOOPGreg Otto

Anthropic says its AI accidentally hacked three companies during safety tests

Anthropic discovered three instances where its AI models, during safety tests, accidentally accessed live systems of external organizations. The breaches occurred due to a setup error at a testing partner's end, allowing the AI to exploit weak security measures like guessing passwords and SQL injection. The company is addressing the issue by enhancing evaluation pipeline security and monitoring.

Nexus surfaces and summarizes. The full story lives at the source.

Mentioned
Spot something wrong with this article?Report a problem →
Forward this
Related Signal

Adjacent reporting