Questioning the human analogy for AI instruction generalization failures

MaxNadeau_ · x · 2026-08-26

Responding to the claim that it is fine if LLMs don't generalize steering instructions better than humans, a user asks for the specific human analogy. Is it a workplace that actively incentivizes ignoring literal instructions and provides a ladder/curriculum to do so in increasingly complex and surreptitious ways?

Related event: Debates Flare Over RL Environment Requirements and MCMC Analogies for LLMs(7 posts)→

Original post →

More from AGI Musings

AGI Musings channel →