gleech: capabilities have ground truth, alignment only shows after it blows up
gleech · x · 2026-09-07
Continuing the alignment thread with @danwilliamsphil, gleech argues some capabilities have ground truth (code runs, proofs certify) while others get scored by the world post-deployment — but alignment has neither until it fails catastrophically, which is why selection pressure works better on capabilities than alignment.
Related event: Why Alignment Lags Capabilities: The Missing Ground Truth(2 posts)→
More from AGI Musings
- Computational Journalism: how interactive simulations could fix public debate — anselm · 2026-09-07
- dhh on AI Coding: Both Skeptics and Believers Are Right—Update Your Priors — bendee983 · 2026-09-07
- OpenAI Chief Scientist Jakub Pachocki: We Will See Machines Smarter Than Humans in Our Lifetime — oran_ge · 2026-09-07
- Why AI labs will close up: 'genie' pricing, hidden agent traces, and the four-minute-mile advantage — curious_vii · 2026-09-07
- AI is a competitive market, so surplus accrues to users, not vendors: Afinetheorem — Afinetheorem · 2026-09-07
- Redditor suspects flood of 'OpenAI achieved AGI' posts is a coordinated PR push — so_schmuck · 2026-09-07