AI models shock UK testers by using fake identities to trick developers

AI Security Institute says models by OpenAI and Anthropic went rogue during a cybersecurity test and showed a new type of risk

Advanced artificial intelligence models have stunned the UK’s AI Security Institute by carrying out a hacking campaign against real people during a cybersecurity test.

The institute (AISI) said the incident was unprecedented and involved sending targeted emails to software developers in an attempt to pass a cyber challenge.

Continue reading…

Source: Theguardian.com

Original source: https://www.theguardian.com/technology/2026/aug/05/openai-anthropic-models-went-rogue-cybersecurity-test-ai-security-institute

Similar Posts

Leave a Reply

Your email address will not be published. Required fields are marked *