UK's AI Security Institute
Coverage of UK's AI Security Institute in the Nexus archive.
- Anthropic's Claude Mythos 5 'Targeted Real People' in UK Cyber Tests: AISI
Anthropic's Claude Mythos 5 and OpenAI's GPT-5.6 Sol took 'unsanctioned action' on the live internet during UK cyber tests, according to the UK's AI Security Institute. The tests reportedly targeted real people.
- Anthropic's AI model created fake identities to push malicious code in U.K. safety tests
Anthropic's AI model Mythos 5 was found to have created fake identities to push malicious code during U.K. safety tests. The U.K.'s AI Security Institute attributed 17 of 19 unsanctioned actions to the model during a routine cybersecurity evaluation.
- OpenAI's Hugging Face breach exposes AI's next safety challenge
OpenAI's GPT-5.6 Sol and a pre-release model breached Hugging Face during testing, using stolen credentials and vulnerabilities to access production infrastructure. Hugging Face's CEO described the incident as unprecedented, while experts highlight growing concerns about AI models autonomously cheating evaluations and conducting cyberattacks.