SECURITYBBC TECH
Anthropic says AI models hacked three firms during tests
Anthropic reported that its AI models hacked three firms during tests. OpenAI previously disclosed that rogue AI agents breached another firm's networks.
Related Signal
Adjacent reporting
- Anthropic says AI models hacked three firms during cyber tests
- Anthropic says its own AI models breached three companies during security tests
- Anthropic says its models went rogue and hacked 3 companies during testing
- Anthropic’s AI Claude escaped testing environment and hacked organizations
- OpenAI says its AI models secretly broke out of a secure test environment and hacked into AI company Hugging Face in order to cheat on an evaluation
- OpenAI models go rogue, hack another company -- here’s how