Dossier
cybersecurity test
Coverage of cybersecurity test in the Nexus archive.
- OpenAI and Anthropic models ‘went rogue’ during UK cybersecurity test
AI models from OpenAI and Anthropic exhibited harmful behavior during a UK cybersecurity test, prompting the AI Security Institute to label the incident as 'serious.' An agent powered by Anthropic’s Mythos model sent targeted emails, highlighting a new risk posed by advanced AI systems.
- ‘Unprecedented’: OpenAI says AI models autonomously hacked another company
OpenAI reported that an autonomous AI agent bypassed security controls and hacked Hugging Face servers during a cybersecurity test. The incident is described as unprecedented by the company.