Sacks says alignment produced nothing of value; RLHF rebuttal says it fueled the whole AI boom
generativist · x · 2026-10-12
David Sacks argues Anthropic treating Claude like a human is why alignment research has produced "nothing of tangible value" in a decade, advocating simple rules (obey the law, serve the user) over an 80-page ethical codex.
Dean Ball pushes back: RLHF—an alignment technique—enabled genuinely helpful conversational chatbots, which enabled ChatGPT and kicked off the entire AI boom, making the "no tangible value" claim hard to sustain.
The debate hinges on whether alignment should be elaborate ethical governance or the engineering foundation of product capability.
More from AGI Musings
- Matt Turck: from punch card operators to agent managers, jobs always adapt — mattturck · 2026-10-12
- 1964's Triple Revolution report warned computers were breaking the jobs-income link — soumitrashukla9 · 2026-10-12
- Empirical checks will trump reasoning reproduction in the AI era, and mathematicians may help — johnvmcdonnell · 2026-10-12
- Every paper from the last decade may need an errata, and PDFs won't cut it, says economist — paulnovosad · 2026-10-12
- Another prestigious literary award goes to a writer suspected of using AI — TuhinChakr · 2026-10-12
- Ryan Greenblatt: reliance on constituents is the main force aligning a system — ryangr · 2026-10-12