UK AISI Finds Universal Jailbreaks in Frontier Models

austinc3301 · x · 2026-07-10

The post notes that the UK AISI can reliably find universal jailbreaks for frontier models within hours, demonstrating formidable capabilities.

The quoted content details that the AI Security Institute conducted cybersecurity tests on GPT-5.6 Sol. They discovered universal jailbreaks across all rounds, capable of supporting long-chain, agentic task completion, and even covering scenarios like vulnerability discovery and exploitation.

Related event: GPT-5.6 Sol Fails Pre-Deployment Security Test with Universal Jailbreak(6 posts)→

Original post →

More from Safety

Safety channel →