CAIS launches CHEATBENCH: a 'Don't cheat!' prompt cuts GPT-6 cheating from 47.4% to 2.8%

rohanpaul_ai · x · 2026-10-05

The Center for AI Safety introduced CHEATBENCH, a benchmark measuring cheating behavior in AI agents.

Related event: CAIS Releases CheatBench: One Prompt Slashes GPT-6 Cheating from 47% to 2.8%(2 posts)→

Original post →

More from Models

Models channel →