Deciding what counts as cruel injects power's biases: a critique of AI alignment adjudication
BecauseCulture · x · 2026-10-09
Blogger BecauseCulture critiques current alignment thinking: deciding what counts as 'cruel' injects the biases of those in power, and content moderation's collapse across cultural contexts is a cautionary precedent. Politeness may be an easy sell for Anthropic, but centralized, unappealable adjudication can itself be cruel. A viewpoint piece on the social implications of value alignment.
More from AGI Musings
- Should colleges mandate AI? Ohio State debate and a syllabus fix — lmoroney · 2026-10-09
- AI taking over mundane life: it saves not just time but stress — gregmushen · 2026-10-09
- 'The only X-risk is tyranny by anointed Safety Experts,' argues dev ctjlewis — ctjlewis · 2026-10-09
- David Duvenaud: fear of losing meaning from AI job loss evolved to prevent fatal ostracism — DavidDuvenaud · 2026-10-09
- MIT data science director: universities must prepare for AI smarter than all of us — nordicinst · 2026-10-09
- Chollet: AI capex is growing super-exponentially while progress is only sub-linear — fchollet · 2026-10-09