Jeff Ladish: concerned people studying rogue AI failures shouldn't all leave the labs
JeffLadish · x · 2026-09-10
Jeff Ladish adds nuance to his argument on lab talent: it's weaker for pretraining team members to stay if they believe catastrophe is likely, but it seems unwise for researchers studying rogue AI failures to all leave, since some crucial research like interpretability can only be done inside labs.
More from AGI Musings
- Anthropic pretraining researcher quits, accusing OpenAI and Anthropic of recklessly racing to self-improving superintelligence — TinfoilTricorn · 2026-09-10
- Researcher Blasts OpenAI Whistleblower Interview: Focus on AI Ethics, Not Alignment — examachine · 2026-09-10
- Researcher calls AI alignment 'safety theater': ethics, not alignment, is the real problem — examachine · 2026-09-10
- Don't buy the 'software engineering is doomed' narrative from AI labs eyeing IPOs — bendee983 · 2026-09-10
- Ex-DeepMind, now Anthropic researcher: no viable scientific plan for recursively self-improving AI risks — harris_edouard · 2026-09-10
- Hugging Face CEO: If AI Risk Is Real, Labs Must Openly Share Models — Cue the Sarcasm — Gradio · 2026-09-10