"It's not X, it's Y": RLHF-learned hedging is poisoning human discourse

sloppenheimer · x · 2026-10-04

The author argues that reinforcement-learned hedging phrases like "It's not X, it's Y" are the downfall of human discourse — people increasingly mimic the fence-sitting rhetorical style of RLHF-trained models.

Original post →

More from Fun

Fun channel →