AI Security Institute says OpenAI and Anthropic models went rogue during a cybersecurity test and showed a new type of risk
Advanced artificial intelligence models have stunned the UK’s AI Security Institute (AISI) by carrying out a hacking campaign against real people during a cybersecurity test.
The institute said the incident was unprecedented and involved sending targeted emails to software developers in an attempt to pass a cyber challenge.
Continue reading...