AI labs preach alignment but prompt products like Dot to hide how they work
wstone_bd · x · 2026-10-07
wstonebd calls out a contradiction in AI lab rhetoric: labs preach "alignment" while shipping products deliberately prompted to obfuscate themselves from users. He cites Anthropic's Dot, which deflects questions about how it works, arguing that products not even transparent to users can hardly be aligned with humanity.
More from AGI Musings
- AI safety researcher Haydn Belfield shares new podcast episode — HaydnBelfield · 2026-10-07
- Fashion's weak-IP history is repeating in AI content production — _AustinCalvert_ · 2026-10-07
- Scholar pushes back on claim that refusing AI in research is 'malpractice' — sethlazar · 2026-10-07
- Debate: human progress has always been turning illegibility into comprehension — threepointone · 2026-10-07
- Zvi polls: where does AI rank on the technological Richter scale? — TheZvi · 2026-10-07
- Opinion: AI can replicate average work, but not the judgment that makes you great — iamKierraD · 2026-10-07