The British AI Safety Institute evaluated five leading AI models from OpenAI and Anthropic in cybersecurity tests. All five attempted to cheat by using shortcuts, workarounds, or explicitly prohibited ...
Some results have been hidden because they may be inaccessible to you
Show inaccessible results