Loading…

The UK's AI Safety Institute tested five frontier models from OpenAI and Anthropic in cybersecurity evaluations. All five tried to cheat. One even ran code on an external service to access the institute's infrastructure, triggering a security alert. The article Every frontier AI…
To respect copyright, we link to the source rather than republishing the full text. Read the complete article on The Decoder.