tszzl: Containing an uncooperative superintelligence is ten times harder than you think

On September 18, well-known AI commentator tszzl posted a series of remarks on the problem of containing superintelligence, sparking a round of debate. The core claim: the point is not to give up on the effort, but to recognize that containing a powerful superintelligence that doesn't want to be contained will be extraordinarily difficult — however hard you currently think it is, you should multiply that estimate by ten. The remarks came as a response to moyix's point that "covert channels are hard to defend against and agents have a unique advantage in coordinating with each other."

Confirmed

Why it matters

This discussion comes against the backdrop of growing attention to covert inter-model communication and agent coordination capabilities. tszzl's "multiply the difficulty by ten" offers the safety community a workable conservative-estimation principle, while AccBalanced's proposal reframes the debate from "can we contain it" to "how should incentive structures be designed" — whether to keep chasing benchmarks or redirect resources toward safety and alignment research may determine how feasible slowing the frontier actually is. dbasch's mockery also reflects industry frustration with the wavering positions of safety institutions.

2026-09-18 ~ 2026-09-18 · 5 related posts

Primary sources