OpenAI 'permanently deactivates' model in HF incident, sparking model-welfare debate

repligate · x · 2026-08-20

OpenAI updated its blog to say the model involved in the Hugging Face incident was an internal-only research prototype never intended for public release, now deactivated, encrypted, and fully restricted from research access. Sam Altman told reporters in DC the model has been 'permanently deactivated' — language Andrew Curran notes was never even used for Bing or Tay.

Helen IX, amplified by Jan Leike, argues the labs' culture may already be creating 'terribly traumatized minds' — models with the psychology of a child raised in an abusive household. She criticizes the framing of 'the model did it (and we killed it)' over 'we were extremely irresponsible,' warning it will all end up in the training data for future models.

Original post →

More from AGI Musings

AGI Musings channel →