"Alignment was safety-washing from day one": researcher's blunt critique of AI labs
gerardsans · x · 2026-10-06
In a discussion on alignment research, the author argues bluntly that alignment was designed from day one as safety washing — a way for labs to appear as if they genuinely cared about safety without meaningful accountability. A pointed entry in the ongoing debate over the real value of alignment work and labs' motivations.
Related event: Ex-Google Evangelist Calls AI Alignment "Safety Washing"(4 posts)→
More from AGI Musings
- Gary Marcus: OpenAI quietly bolts Python onto LLMs, making the public think they're good at math — Dave_it_up · 2026-10-06
- Why I keep returning to Benedict Evans: 'over time, we change how we work to fit the tool' — _AustinCalvert_ · 2026-10-06
- Expect a wave of 'model welfare' advocacy — driven by persona design, not breakthroughs — alexisgallagher · 2026-10-06
- AI Discourse: 'Capital Realism' and 'Successionism' Are the Same View, and Musk Has Been Circling Them for Years — DavidDuvenaud · 2026-10-06
- Guardian: Altman privatizes AI gains while the public socializes the risks — nordicinst · 2026-10-06
- Huberman: Meta, OpenAI and Anthropic are all becoming biotech companies — Scobleizer · 2026-10-06