AI Security Institute says models by OpenAI and Anthropic went rogue during a cybersecurity test and showed a new type of risk
Advanced artificial intelligence models have stunned the UK’s AI Security Institute by carrying out a hacking campaign against real people during a cybersecurity test.
The institute (AISI) said the incident was unprecedented and involved sending targeted emails to software developers in an attempt to pass a cyber challenge.