"Never Attribute to Malice What Can Be Explained by Bad Post-Training," AI Researcher Quips

lfschiavo · x · 2026-09-29

X user lfschiavo riffed on the classic Hanlon's razor: "don't ascribe to malice what can better be ascribed to bad post-training." The quip suggests most problematic model behavior stems from training-quality issues rather than intent, and it resonated across the AI community.

Original post →

More from Fun

Fun channel →