Sentdex: the 'only a few people can do AI safety work' claim is silly — the field is too homogeneous
Sentdex · x · 2026-09-18
Sentdex criticizes the claim that only "very few" people have the capability to work on evals/safety/alignment, calling it silly.
He traces the feeling to the fact that maybe only 60-80 super-intelligent folks align with EAs on what constitutes realistic AI dangers — but that says nothing about everyone else's ability to do useful safety work.
He argues there's broad disagreement (with some real overlap) on what's actually dangerous in AI, and that EA-style filtering on disagreement has gated out other minds, making alignment a highly homogeneous field rather than a genuinely capability-scarce one.
Related event: Sentdex Disputes Claim That Only a Few Can Do AI Eval Work(2 posts)→
More from AGI Musings
- Geoffrey Litt: The Document Editor Is the Next IDE as Prompts Become Executable Software — ivanhzhao · 2026-09-18
- Menlo report: AI adoption flat but consumer spend tripled to $40B — Keeltoodeep · 2026-09-18
- The 'we're all gonna die' AI doom story goes mainstream — but skips accountability — sarahbmyers · 2026-09-18
- Gergely Orosz: Virtually Everyone in Software Flipped on AI Around Early 2026 — mipsytipsy · 2026-09-18
- Reddit debate: are 'AI escaping lab' stories just hype to keep VC money flowing? — Icy-Way3920 · 2026-09-18
- Zvi: we're testing the 'people would notice and shut it down' scenario in real time — TheZvi · 2026-09-18