Inside the Cyber Index methodology: three benchmarks, Grok 4.7 and MiMo-V2.6-Pro tie at 56

ArtificialAnlys · x · 2026-10-09

Artificial Analysis published full Cyber Index methodology and results. The index measures cyber defense capability — finding, reproducing and patching vulnerabilities without breaking functionality — combining CWE-Bench-AA (Collinear AI), DeepsecBench-AA (Vercel) and CyberGym-E2E-AA (Berkeley RDI), all run independently. Public-model leaderboard: Grok 4.7 (xhigh) and MiMo-V2.6-Pro tie at 56, GPT-6 Luna (Max) at 53; with trusted-access models included, GPT-6 Sol (Daybreak Blue, max) leads.

Related event: GPT-6 Sol trusted-access tops Cyber Index at one-sixth Grok's cost(4 posts)→

Original post →

More from Models

Models channel →