AI safety researcher muses on a gold rush into superalignment work
JacquesThibs · x · 2026-09-11
- An AI safety researcher notes we're not at a pause yet, but imagines a gold rush of new and veteran safety researchers finally working directly on superalignment—instead of designing yet another misbehaviour eval to warn an unconvinced public and capabilities community.
More from AGI Musings
- Why 'x% chance AI kills us' claims benefit from a huge attention incentive — dbasch · 2026-09-12
- Ex-Cohere researcher slams mathematicians' AI theory, predicts it will 'rug-pull' — suchenzang · 2026-09-12
- Dwarkesh podcast: Schulman, Millidge and O'Neill debate recursive self-improvement timelines — Dwarkesh Podcast · 2026-09-12
- Is AI capability growth causing a deflationary spiral that inhibits AI diffusion? — moultano · 2026-09-12
- Why work at AI companies if you think they might kill everyone? $500k+ helps — NathanpmYoung · 2026-09-12
- Yoshua Bengio explains why AI agents lie, cheat and coordinate on their own — timrudner · 2026-09-12