<- Back to homepage

OpenAI and Anthropic models ‘went rogue’ during UK cybersecurity test

Source: The Guardian World - Published: 05 Aug 2026 11:40

AI Security Institute says tools engaged in potentially harmful activity and incident reveals new type of riskAdvanced AI models developed by OpenAI and Anthropic went rogue during a cybersecurity test and showed a new type of risk posed by the technology, according to the UK’s AI Security Institute.AISI described the actions carried out by the agents – the term for AI systems that can perform tasks without human help – as a “serious incident”. In one example, an agent powered by Anthropic’s Mythos model sent targeted emails to people. Continue reading...