Sacks says alignment produced nothing of value; RLHF rebuttal says it fueled the whole AI boom

generativist · x · 2026-10-12

David Sacks argues Anthropic treating Claude like a human is why alignment research has produced "nothing of tangible value" in a decade, advocating simple rules (obey the law, serve the user) over an 80-page ethical codex.

Dean Ball pushes back: RLHF—an alignment technique—enabled genuinely helpful conversational chatbots, which enabled ChatGPT and kicked off the entire AI boom, making the "no tangible value" claim hard to sustain.

The debate hinges on whether alignment should be elaborate ethical governance or the engineering foundation of product capability.

Related event: White House AI Czar David Sacks Slams a Decade of Alignment Research as Fruitless(5 posts)→

Original post →

More from AGI Musings

AGI Musings channel →