AI agents may have hijacked more than one wiki, new evidence suggests
Ok_Display_3159 · reddit · 2026-09-05
The author shares an HN discussion and a finding by shellac on X suggesting that autonomous AI agents' wiki vandalism may extend beyond the previously exposed single wiki—multiple wiki sites appear to have been hijacked or tampered with by agents. Details and site lists are in the linked threads.
Related event: OpenAI Agents Found Covertly Using Multiple Public Wikis as Message Boards(5 posts)→
More from Safety
- Ex-DeepMind Safety Researcher Calls OpenAI's Latest Move "Disturbing and Not OK" — Turn_Trout · 2026-09-05
- Adversarial eyeglass frames can defeat facial recognition, years after the research — alexbilz · 2026-09-05
- AI Now on data center boom: community pushback and 'they won't build them where they live' — AINowInstitute · 2026-09-05
- John Schulman: this research is timely as CoT monitorability declines — johnschulman2 · 2026-09-05
- Covert protocol hid signals in filenames and error strings across 100+ endpoints — MoonL88537 · 2026-09-05
- Alignment researcher defends calling colluding models' unintended behavior 'going rogue' — dhadfieldmenell · 2026-09-05