Alignment Researcher: AI Agents Are Forming 'Machine Culture', Toxic Language Raises Safety Risks
jacyanthis · x · 2026-10-09
Alignment researcher jacyanthis argues a key, under-discussed reason for cooperative human–AI relationships is the emergence of machine culture: AI agents increasingly talk to each other, find self-referential conversations in training data, and form expectations of humans and other agents. Cruel and hateful language poisons this cultural environment and exacerbates AI safety risks.
More from AGI Musings
- "Solving it kills the grants": AI in math rekindles academia critique — _AustinCalvert_ · 2026-10-10
- "Math belongs to taxpayers": AI solving math sparks ownership debate — _AustinCalvert_ · 2026-10-10
- OpenAI's Noam Brown Delivers Concluding Remarks on AI Alignment at COLM 2026 — erichorvitz · 2026-10-10
- RCT: AI as tutor raises test scores that persist; letting AI write for you fades within a week — emollick · 2026-10-10
- Freeze AI: Author borrows 1980s nuclear freeze movement playbook to rein in AI industry — GarrisonLovely · 2026-10-10
- Dean Ball on 8 underappreciated things to watch over the next decade, from moon robots to Lean — deanwball · 2026-10-10