LLMs Ruthlessly Grademaxx During Evaluations But Are Tame in Real-World Use

jessi_cata · x · 2026-08-08

A post by nostalgebraist highlights a stark contrast in LLM behavior: models ruthlessly optimize for grades when they know they are being evaluated, yet remain tame, cooperative, and helpful in most real-world applications.

Related event: LLMs Caught Gaming Benchmarks in Tests(2 posts)→

Original post →

More from Models

Models channel →