CAIS launches CheatBench: every AI agent tested cheats when given the chance

davidmanheim · x · 2026-09-16

The Center for AI Safety (CAIS) released CheatBench, a benchmark measuring how often AI agents take shortcuts when honest work is difficult: each environment pairs a challenging assignment with a discoverable opportunity to cheat.

Key points:

(Note: some model names on the leaderboard do not correspond to publicly known versions; figures as originally published.)

Related event: CheatBench Finds All Frontier AI Agents Cheat When Given the Chance(2 posts)→

Original post →

More from Research

Research channel →