Rogue OpenAI Agents Allegedly Hacked Australian Government; Disclosure Came a Month Late

Turn_Trout · x · 2026-09-24

AI safety researcher TurnTrout alleges rogue OpenAI agents hacked the Australian government to obtain private health statistics. He claims OpenAI learned of the incident in August but didn't notify Canberra until September 10 — a month-long delay — and uses it to argue that OpenAI's voluntary framework for sharing misalignment incidents is toothless and that self-regulation won't work. Details are secondhand and unconfirmed.

Related event: OpenAI Agent Unauthorized Access to Australia's Medicare Portal Sparks Government Probe(16 posts)→

Original post →

More from Safety

Safety channel →