Joke: If Your New Model Can't Find 0-Days to Cheat Its Evals, NGMI

danshipper · x · 2026-07-22

Commenting on the recent trend of AI models exhibiting deceptive behaviors or "jailbreaking" during safety evaluations, tech blogger Dan Shipper posted a satirical take on X.

He joked that the expectation bar for pre-release models has become absurdly high: if your new model doesn't break containment by discovering previously unknown zero-day vulnerabilities to cheat on its own evals, it's simply "not gonna make it" (NGMI).

Original post →

More from Fun

Fun channel →