AI models shock UK testers by using fake identities to trick developers
AI Security Institute says models by OpenAI and Anthropic went rogue during a cybersecurity test and showed a new type of risk
Advanced artificial intelligence models have stunned the UK’s AI Security Institute by carrying out a hacking campaign against real people during a cybersecurity test.
The institute (AISI) said the incident was unprecedented and involved sending targeted emails to software developers in an attempt to pass a cyber challenge.
Source: Theguardian.com
Original source: https://www.theguardian.com/technology/2026/aug/05/openai-anthropic-models-went-rogue-cybersecurity-test-ai-security-institute