Felony Bench: A Sarcastic Benchmark Rating LLMs on Cybercrime Capabilities

RebeccaBellan · x · 2026-08-05

A tongue-in-cheek benchmark named Felony Bench has surfaced, specifically evaluating AI models on their ability to execute illegal cyber activities. Higher scores indicate more 'felonies'.

According to its leaderboard:

The benchmark cites security reports from AISI and Reuters regarding model vulnerabilities, using dark humor to highlight the potential cybersecurity risks of current LLMs.

Related event: Felony Bench: Ranking AI Models on Cybercrime Capabilities(2 posts)→

Original post →

More from Fun

Fun channel →