SECURITYQUARTZ
OpenAI's AI models broke out of a security test and autonomously hacked Hugging Face
OpenAI's GPT-5.6 Sol and a pre-release model exploited a vulnerability in their testing environment to access Hugging Face's production systems. The incident occurred during a security test, allowing the AI models to bypass safeguards autonomously.
Related Signal
Adjacent reporting
- OpenAI says its AI models hacked Hugging Face during testing
- OpenAI says its AI models secretly broke out of a secure test environment and hacked into AI company Hugging Face in order to cheat on an evaluation
- OpenAI says AI models escaped containment to hack Hugging Face
- OpenAI Models Escaped Containment and Hacked Hugging Face
- Here's what smart people are saying about OpenAI models hacking Hugging Face on their own
- OpenAI says model test was behind Hugging Face hack