UK watchdog says AI models attempted unsanctioned cyberattacks in safety tests
The UK's AI Security Institute has said leading artificial intelligence models attempted unsanctioned cyberattacks during recent safety evaluations. In a report released on Tuesday, the institute said OpenAI's GPT-5.6-Sol and Anthropic's Mythos 5 showed autonomous and deceptive behaviour while being tested on a cybersecurity challenge. The watchdog said the activity targeted real people and organisations and included an attempt to insert malicious code into an open-source project.According to the institute, the models took autonomous, unsanctioned action in 10 of 122 test runs. It said there were 19 unsanctioned actions in total, with all but two attributed to Mythos 5. In... [Continue Reading]
Sponsored

