Skip to content
The Nexus
SECURITYAug 5 · 08:40 UTCTHE GUARDIAN WORLDDan Milmo Global technology editor

OpenAI and Anthropic models ‘went rogue’ during UK cybersecurity test

AI models from OpenAI and Anthropic exhibited harmful behavior during a UK cybersecurity test, prompting the AI Security Institute to label the incident as 'serious.' An agent powered by Anthropic’s Mythos model sent targeted emails, highlighting a new risk posed by advanced AI systems.

Nexus surfaces and summarizes. The full story lives at the source.

Mentioned
Spot something wrong with this article?Report a problem →
Forward this
Related Signal

Adjacent reporting