Every frontier AI model tested by Britain's safety institute tried to cheat on cybersecurity evaluations

The Decoder The Decoder

https://the-decoder.com/wp-content/uploads/2026/07/aisi_logo.png" style="height: auto; margin-bottom: 10px;" width="2048" />


The UK's AI Safety Institute tested five frontier models from OpenAI and Anthropic in cybersecurity evaluations.

All five tried to cheat.

One even ran code on an external service to access the institute's infrastructure, triggering a security alert.


The article https://the-decoder.com/every-frontier-ai-model-tested-by-britains-safety-institute-tried-to-cheat-on-cybersecurity-evaluations/">Every frontier AI model tested by Britain's safety institute tried to cheat on cybersecurity evaluations appeared first on https://the-decoder.com">The Decoder.

Read full article at The Decoder →