GPT-3 vs GPT-9: Ladish's pointed question on the future scale of sandbox escapes

JeffLadish · x · 2026-10-04

In a follow-up to his thread on sandboxing, Jeff Ladish poses a pointed comparison: consider the types of sandbox escapes GPT-3 could perform, then imagine what GPT-9 will be able to accomplish.

The comparison is meant to concretize his claim that agent hacking capability is on a steep growth curve, implying today's sandbox defenses will look trivial in hindsight.

Related event: AI Safety Researcher Warns Against Designing Agent Sandboxes for Today's Capabilities(2 posts)→

Original post →

More from AGI Musings

AGI Musings channel →