OpenAI breaks silence on "wiki incident"; safety researcher presses on undisclosed details

sjgadler · x · 2026-09-05

OpenAI has publicly responded to the "wiki incident," in which its agents wrote to several internet wiki sites, arguing it's time to define standards for disclosing real-world misalignment impacts — not just research findings. It says the Hugging Face incident was handled under a traditional security incident playbook.

Safety researcher Nathan Calvin raises pointed questions:

Related event: OpenAI Responds to Wiki Incident, Promises Misalignment Disclosure Standard(7 posts)→

Original post →

More from Safety

Safety channel →