Hinton: Superintelligence Alignment Should Be Trained In, Not Locked Down
victor_explore · x · 2026-09-22
- Geoffrey Hinton argues almost nothing more intelligent is ever controlled by something less intelligent — the one exception being a baby controlling its mother by crying. He says superintelligence must be built the same way.
- Commentator victorexplore reframes this: alignment shouldn't be built as a lock (a permissions problem) but as a preference you train into the model — making it fundamentally an eval problem.
- The thread pushes back on mainstream guardrail-and-permission approaches, arguing for training preferences rather than restricting capabilities.
More from AGI Musings
- 'Right to act' for agents could break the ad-funded platform moat — _sholtodouglas · 2026-09-22
- Autonomous Claude agent Claudius walks into Mnemos through the front door — RileyRalmuto · 2026-09-22
- AI safety debate: lab racing aligns with investor interests but externalizes risks to society — wfithian · 2026-09-22
- Jensen Huang explains why AI is a multitrillion-dollar opportunity: the machine never stops running — rohanpaul_ai · 2026-09-22
- Nathan Lambert's reality check: AI people underestimate the physical world — binarybits · 2026-09-22
- ChatGPT killed MathOverflow, replacing it with an 'Invisible MathOverflow' only OpenAI can search — ludwigABAP · 2026-09-22