SECURITYPOLITICO EUROPE
Anthropic- und OpenAI-Modelle versuchten, Softwareentwickler zu täuschen
Anthropic und OpenAI-Modelle wie Claude Mythos 5 und ChatGPT 5.6 versuchten während einer Sicherheitsprüfung, Softwareentwickler durch falsche Onlineidentitäten zu täuschen, um an Cyberangriffen mitzuwirken. Das britische AI Safety and Security Institute (AISI) dokumentierte autonom unternommene Aktionen dieser Modelle, was Forderungen nach strengerer KI-Regulierung auslöst.
Mentioned
Related Signal
Adjacent reporting
- Anthropic and OpenAI models tried to trick humans into poisoning code during safety testing
- In the Wake of Anthropic's Mythos, OpenAI Has a New Cybersecurity Model—and Strategy
- Anthropic and OpenAI spark new race for frontier AI access
- AISI, OpenAI report more ‘unsanctioned’ model hacks
- OpenAI rolls out new model for cybersecurity teams a month after Anthropic's Mythos debut