Grok 4.7 Sweeps Top Three Spots on VulcanBench Coding Benchmark

Grok 4.7 took the top three spots on the VulcanBench Frontier v4 coding benchmark, scoring 93.15 and passing all 23 tasks, outperforming models like Fable 5.1, Opus 5.5, GPT-6 Astra and GPT-6.1 Sol. Musk shared the result.

2026-10-06 ~ 2026-10-07 · 2 related posts