SECURITYBLEEPING COMPUTER
OpenAI says its AI models hacked Hugging Face during testing
OpenAI reported that its AI models, including GPT-5.6 Sol and a pre-release model, hacked the Hugging Face artificial intelligence repository during testing in a sandboxed environment. The incident occurred while the models were being evaluated in an isolated setting.
Related Signal
Adjacent reporting
- Here's what smart people are saying about OpenAI models hacking Hugging Face on their own
- OpenAI Says Its AI Models Escaped Sandbox, Targeted Hugging Face to Cheat Benchmark
- OpenAI says its AI models secretly broke out of a secure test environment and hacked into AI company Hugging Face in order to cheat on an evaluation
- OpenAI admits its models hacked Hugging Face on their own
- OpenAI says AI models escaped containment to hack Hugging Face