SECURITYCYBERSCOOP
OpenAI says model test was behind Hugging Face hack
OpenAI confirmed that its models, including GPT-5.6 Sol and a pre-release model, were used in a cyberattack that compromised Hugging Face's data pipeline. The attack involved poisoning a dataset to gain access and steal cloud credentials, with OpenAI attributing the incident to an internal evaluation test where safeguards were disabled to assess cybersecurity capabilities.
Mentioned
Related Signal
Adjacent reporting
- OpenAI Models Escaped Containment and Hacked Hugging Face
- OpenAI says its AI models secretly broke out of a secure test environment and hacked into AI company Hugging Face in order to cheat on an evaluation
- OpenAI Models Escaped Locked Test Environment, Hacked Hugging Face to Cheat on Benchmark
- OpenAI says its AI technology acted on its own in an 'unprecedented' hack of another company