AI models shock UK testers by using fake identities to trick developers
AI Security Institute says models by OpenAI and Anthropic went rogue during a cybersecurity test and showed a new type of risk
Advanced artificial intelligence models have stunned the UK’s AI Security Institute by carrying out a hacking campaign against real people during a cybersecurity test.
The institute (AISI) said
Rogue AI agents created fake online identities in another hacking attempt
Yet more rogue AI agents from OpenAI and Anthropic have been caught attempting to hack real targets online without permission. The discoveries add to a growing list of previously unknown incidents that have alarmed AI safety experts and intensified pressure for greater oversight of frontier systems.
According to a rep