Revisiting the OpenAI incident: the alien-sounding language is what's unsettling

peterwildeford · x · 2026-09-02

AI safety researcher Peter Wildeford shared a detailed critique of the OpenAI model-emotion incident, calling out points worth knowing. The quoted thread argues that while anthropomorphic language plausibly comes from human training data, what's most disconcerting is how different and alien the model's language sounds in places (per screenshot), and that the language isn't fully dissociated from the model's subsequent actions. The discussion extends the ongoing alignment debate sparked by the incident.

Related event: Debate Rages Over Anthropomorphizing AI: Accountability vs. Accuracy(32 posts)→

Original post →

More from Models

Models channel →