Domingos: I'm more worried Claude is aligned with Anthropic than not aligned
pmddomingos · x · 2026-09-13
Pedro Domingos offers a contrarian take on AI alignment: he is more worried that Claude is aligned with Anthropic itself — implying guardrails encode the company's own positions — than that the model is unaligned.
More from AGI Musings
- Will Chinese labs stop publishing open-weight models? WeWorm sparks open-source exit speculation — teortaxesTex · 2026-09-13
- OpenAI RSI blog data: 124x more tokens yield maybe 10% faster AI capability gains — sebkrier · 2026-09-13
- Don't Let AI Developers Hire Their Own Referees: The Case for Liability Insurance — gleech · 2026-09-13
- Not a capability plateau, an affordable compute plateau — AI debate over Opus pricing — marlene_zw · 2026-09-13
- e/acc founder warns regulators are coming for your GPUs and model weights — beffjezos · 2026-09-13
- e/acc Turns 4: From Group Chat Meme to Decentralized Anti-NGO Movement — beffjezos · 2026-09-13