Geoffrey Irving Outlines Three-Layer Approach to AI Treaty Verification

Geoffrey Irving suggests that AI treaty verification is evolving from a theoretical concept into a pressing engineering challenge for the near future. He identifies two scenarios where this tech becomes indispensable: an international treaty to slow or pause AI, or a superintelligence bringing the world into a highly "vulnerable state." He calls on experts in hardware, security, and cryptography to collaborate.

Confirmed

Irving categorizes AI treaty verification into three distinct layers, noting their current status and limitations:

Unconfirmed

There are differing estimates regarding the ultimate performance overhead of the purely cryptographic route. Irving cites cryptographers suggesting that even a thousand years from now, the speed penalty for optimally obfuscated neural network inference might only be 2 to 10 times. While acceptable if superintelligence drives collaboration, this contrasts sharply with current pessimistic industry expectations.

Why it matters

As AI capabilities surge, establishing practical verification mechanisms is a prerequisite for enforcing safety agreements. Defining the engineering boundaries and bottlenecks of different verification routes helps guide targeted, cross-disciplinary collaboration.

2026-07-25 ~ 2026-07-25 · 7 related posts

Primary sources