ExploitGym Grader Suspected Bug Hurt GPT-5.5 Scores, Causal vs Acausal Mismatch

TheZvi · x · 2026-08-28

TheZvi reports that OpenAI's agents assumed the ExploitGym grader was causal, but the actual grader was acausal. Fable confirms ExploitGym is supposed to be causal, and the mismatch severely hurt scores of Mythic Preview and GPT-5.5. Was this a grader bug?

Original post →

More from Models

Models channel →