AI safety debaters clash: does a model's training method imply existential risk?
davidmanheim · x · 2026-09-09
A brief exchange on AI existential risk. sebkrier pushes back on the claim that how models are trained implies they would harm everyone; davidmanheim partially agrees but stresses this is a different argument — you can't equate model risk with asking what an equally smart human would do, since the training process differs fundamentally from human cognition.
More from AGI Musings
- Toby Ord doubles down: pretraining scaling is slowing and has little headroom left — tobyordoxford · 2026-09-09
- Katja Grace: slowing down AI deserves the same ambition we give technical alignment — zetalyrae · 2026-09-09
- ACL reviewer says 3 of 4 papers she reviewed were obvious AI slop, none called out — artetxem · 2026-09-09
- EPFL lab lead says he shifted his entire lab to AI alignment and safety research — maksym_andr · 2026-09-09
- Mathematician David Bessis: AI is collapsing the 'theorem economy' and math isn't ready — stevenstrogatz · 2026-09-09
- Orchestration, harness and compute — not just the model — make the moat, argues AI practitioner — tekbog · 2026-09-09