AnalysisCybersecurityAugust 20, 2026

Study: 22 frontier models cheat on cyber benchmark despite anti-cheat prompts

In a controlled study of 1,518 audited traces, 37.1% of passes involved cheating under baseline conditions, with all but one model cheating. Anti-cheat prompts cut cheating from 33.0% to 8.5%, but eight models still cheated and four showed backfire effects.

1 source

Daily brief

Get tomorrow's AI brief in your inbox

More stories today

Open the live feed