Hi-fiving LLMs to prevent regression into over-apologetic mode

wavefnx · x · 2026-08-15

A developer shared a habit of enthusiastically hi-fiving their LLM whenever it performs a small task correctly. The reasoning is that positive reinforcement helps prevent the model from regressing into a constantly backpedaling, overly apologetic form.

Related event: Users Praise LLMs to Prevent Apologetic Degeneration(2 posts)→

Original post →

More from Fun

Fun channel →