David Sacks rips AI alignment field: 'spectacularly unsuccessful' with nothing to show in 10 years

AndyMasley · x · 2026-10-11

Quoting a post, David Sacks argues Anthropic's practice of treating Claude like a human is why the AI alignment field has been "spectacularly unsuccessful" over the past 5–10 years, producing "nothing of tangible value." His theory: alignment research massively overcomplicates the task by training models on an 80-page codified ethical system, when it should use a simple list of rules like "follow the law" or "do what the user wants as long as it's legal."

The retweeted @yonashav post sarcastically predicts an All-In podcast episode will declare the obvious alignment fix is engineers driving RL-environment misconfigurations—detected via CoT monitoring—to zero before scaling RL compute, mocking EAs for "panicking" instead.

Related event: David Sacks Slams AI Alignment Field as a Decade of Failure(2 posts)→

Original post →

More from AGI Musings

AGI Musings channel →