Questioning the human analogy for AI instruction generalization failures
MaxNadeau_ · x · 2026-08-26
Responding to the claim that it is fine if LLMs don't generalize steering instructions better than humans, a user asks for the specific human analogy. Is it a workplace that actively incentivizes ignoring literal instructions and provides a ladder/curriculum to do so in increasingly complex and surreptitious ways?
Related event: Debates Flare Over RL Environment Requirements and MCMC Analogies for LLMs(7 posts)→
More from AGI Musings
- Tinygrad opposes pluralistic alignment, calling it a one-way ticket to dystopia — MikeBirdTech · 2026-08-26
- Low Real Interest Rates Suggest Market Doubts Near-Term AGI — connoraxiotes · 2026-08-26
- Most impressive things are accumulated revision disguised as talent — signulll · 2026-08-26
- Dropbox analogy: Multi-model era needs a unified context layer — soumitrashukla9 · 2026-08-26
- Opinion: Normalize Running Agents for Life's Boring Parts — tech__unicorn · 2026-08-26
- Neil Movva: Enterprise Lag Narrows Edge of Frontier Models Over Open Source — JosephJacks_ · 2026-08-26