Domingos: AI Alignment Community Unknowingly Building a Censor's Toolkit
pmddomingos · x · 2026-07-06
Pedro Domingos cited an ICML 2026 paper, pointing out that the AI alignment community is unwittingly building a "censor's toolkit." This has sparked criticism and ethical debates over whether alignment work could be repurposed for content moderation. It represents an ongoing debate in the value alignment space.
More from AGI Musings
- Can closed frontier models support CBRN defense, or do we need open ones too? — olcan · 2026-07-27
- Steven Pinker says intelligence is bounded, not a limitless scalar like height — SydSteyerhart · 2026-07-27
- A poster imagines AI evolving into synthetic, collective, and beyond-intelligence systems — CurieuxExplorer · 2026-07-27
- AI surpassing human intelligence will be noticeable, as Claude's cryptic output proves — zetalyrae · 2026-07-27
- AI lab staff have gone strangely quiet about next-year capability predictions — ChrisGPT · 2026-07-27
- AI could erode science by flooding research with credible slop — rbhar90 · 2026-07-27