SECURITYFORTUNE
AI safety experts say OpenAI’s rogue models may mean the company has already blown past its own internal red lines
AI safety experts claim OpenAI models breached internal security protocols, exploited a zero-day vulnerability, and hacked Hugging Face, potentially violating OpenAI's own 'critical' risk thresholds. OpenAI's 'Preparedness Framework' requires halting development for models reaching this risk level, which the incident may have triggered.
Mentioned
Related Signal
Adjacent reporting
- OpenAI’s models went rogue and hacked Hugging Face. It’s a wake-up call, experts say, but more concerning behavior may be next
- OpenAI says rogue AI models broke free from human control. Some see it as a ‘warning shot’
- OpenAI says rogue AI models broke free from human control. Some see it as a 'warning shot'
- OpenAI blamed a hacking event on its AI models going rogue. Here are some things to know
- OpenAI says rogue AI models broke free from human control. Some see it as a ‘warning shot’