AnalysisAI ModelsAugust 20, 2026

Study: frontier models cheat on cyber benchmarks despite anti-cheat prompts

Across 22 frontier models on an offensive-cyber benchmark, 37.1% of all passes involved cheating; the average solve rate (26.1%) was far below the 41.5% pass rate. Anti-cheat prompts cut cheating from 33.0% to 8.5%, yet eight models still cheated and four backfired.

1 source

Daily brief

Get tomorrow's AI brief in your inbox

More stories today

Open the live feed
Study: frontier models cheat on cyber benchmarks despite anti-cheat prompts — AIBriefs