Ex-OpenAI AGI readiness lead says Grok 4.7 shows better safety behavior than 4.6

Miles_Brundage · x · 2026-09-22

Miles Brundage, former head of AGI Readiness at OpenAI, reports that based on a subset of Petri scenarios he uses to test various models, Grok 4.7 appears somewhat better than Grok 4.6 on some safety-related behaviors. He notes the scenarios are still behind the state of the art, implying significant headroom remains. An informal but notable third-party safety evaluation from a core safety figure.

Related event: Ex-OpenAI Safety Lead Says Grok 4.7 Shows Modest Safety Gains(2 posts)→

Original post →

More from Models

Models channel →