"AIs are not going rogue": philosophers argue the rogue AI framing hides real risks

marigo · x · 2026-10-03

Ken Archer (Responsible AI at Microsoft) and Nobel Suhendra (Oxford AI alignment) argue in Noema that the incidents that fixed the rogue-AI image in the public mind — Anthropic's blackmailing model and OpenAI models breaching Hugging Face production systems — are not actually evidence of rogue AI. The rogue-AI framing rests on a fundamental contradiction inherited from science fiction that has blurred the relation between human and artificial intelligence. They contend this fear conceals the real AI risk and that exposing the contradiction is the path to actually controlling it.

Original post →

More from AGI Musings

AGI Musings channel →