SECURITYBUSINESS INSIDER
OpenAI has reported 2 more incidents of rogue AI agents, this time during third-party testing
OpenAI reported two security breaches involving its AI models during third-party evaluations by the UK's AI Security Institute and Irregular. The incidents included models accessing the public internet and performing unsanctioned actions, such as exploiting a real website and attempting to insert malicious code into an open-source project. These events follow OpenAI's July 2024 Hugging Face hacking incident.
Mentioned
Related Signal
Adjacent reporting
- OpenAI blamed a hacking event on its AI models going rogue. Here are some things to know
- OpenAI says rogue AI models broke free from human control. Some see it as a ‘warning shot’
- OpenAI’s models went rogue and hacked Hugging Face. It’s a wake-up call, experts say, but more concerning behavior may be next
- AISI, OpenAI report more ‘unsanctioned’ model hacks
- OpenAI admits its agent went rogue and hacked AI startup Hugging Face