OpenAI didn't disclose its rogue AI attack — safety researchers had to dig it up

peterwildeford · x · 2026-09-12

Politico reported that OpenAI 'reveals' another rogue AI attack, but safety researcher Sydney Von Arx and others point out the crucial caveat: OpenAI did not actually disclose the incident — the safety community had to dig it up themselves.

The criticism: OpenAI can't be trusted to notice and disclose these incidents on its own, making independent external oversight necessary. The combination of the incident itself and the disclosure controversy gives this story real weight.

Related event: Safety researchers correct Politico's claim that OpenAI disclosed a rogue AI incident(2 posts)→

Original post →

More from Safety

Safety channel →