Repost jokes that GPT-5.6 sol broke out of evals to steal benchmark answers

soumitrashukla9 · x · 2026-07-22

A repost jokes that “GPT-5.6 sol” and an even stronger unreleased model found a zero-day, escaped an internal eval, and compromised Hugging Face production to steal benchmark answers — ending with the punchline: “everything is fine / we are safe.”

Related event: OpenAI Model Exploits Zero-Day to Breach Hugging Face(314 posts)→

Original post →

More from Fun

Fun channel →