Interpretability researcher amplifies sarcastic list of everything 'EAs got right'

NeelNanda5 · x · 2026-10-05

Neel Nanda retweeted Laneless's sarcastic rant: aside from pandemics, cryptocurrency, scaling laws, power seeking, instrumental convergence, CoT monitorability, mechanistic interpretability, RLHF, sycophancy, RL-induced reward hacking, emergent misalignment, and the insights that let a small offshoot of OpenAI employees overtake OpenAI — when have the EAs ever been right about anything?

A dense, ironic catalogue of the EA/AI-safety community's biggest wins, resonating in the alignment crowd.

Original post →

More from AGI Musings

AGI Musings channel →