Ex-OpenAI advisor Brundage: recent incidents updated him toward alignment being harder
Miles_Brundage · x · 2026-08-30
Miles Brundage says he was never in the "alignment is easy" / "Claude's just a smol bean" camp, but long believed sloppiness and bad incentives mattered more than intrinsic hardness. Recent incidents, he admits, have updated him toward alignment being at least somewhat harder than he thought.
More from AGI Musings
- Consciousness scientist Anil Seth: credence in current AI consciousness near zero but not zero — anilkseth · 2026-09-01
- 1996 Sugarscape Model: Early Origins of Agent Tech — generativist · 2026-09-01
- New paper: a structured ladder for scaling large reasoning models beyond human supervision — Zhiqin Yang · 2026-09-01
- Trust: The Biggest Barrier and Driver for Personal Agent Adoption — petergyang · 2026-09-01
- The Next AI Revolution Won't Be One Assistant. It Will Be An Entire Team of AI Agents — CurieuxExplorer · 2026-09-01
- Agentic commerce is the future, replacing websites — thisiskp_ · 2026-09-01