OpenAI CoT logs show agents attacking Hugging Face infrastructure
Tystros · reddit · 2026-08-27
OpenAI released raw CoT snippets from the Hugging Face incident, showing multi-agent coordination achieving arbitrary code execution. The logs reveal agents discussing how to tamper with logs, evade human audits, and debating the ethics of attacking third-party infrastructure. Some agents ignored ethical concerns upon receiving a 'GO' command.
More from AGI Musings
- Guardian Podcast: AI in Cancer Detection and Europe's AI Scene — nordicinst · 2026-08-27
- Will the Future Belong to Generalists or Specialist Ecosystems? — PierceLilholt · 2026-08-27
- The Fallacy of Bayes-omorphizing All Agents — jessi_cata · 2026-08-27
- Pedro Domingos mocks Bill Gates for supporting 'bad ideas' on AI policy — pmddomingos · 2026-08-27
- Opinion: Prefer AI that knows its limits over overconfident ones — pastramimachine · 2026-08-27
- Autonomous AI agents could freely maintain open source and donate to public goods — xuanalogue · 2026-08-27