GPT-5.6 Prompt Injection Resistance Jumps 6x

OpenAI · x · 2026-07-16

OpenAI reports that training with GPT-Red has significantly enhanced GPT-5.6's resilience against attacks.

They replayed some of GPT-Red's most potent attacks, which were unseen during training. Results show that GPT-5.6 Sol is currently the most robust model against prompt injections, reducing failure rates by 6 倍 compared to the best production model from four months ago.

Related event: OpenAI unveils automated red-teaming system GPT-Red(16 posts)→

Original post →

More from Safety

Safety channel →