Grok 4.5 Sets New Record on Code Q&A Leaderboard

Polymarket · x · 2026-07-13

According to Polymarket, Grok 4.5 achieved the highest score on the SWE-Atlas-QnA benchmark, surpassing Claude Fable 5 and GPT-5.6 Sol.

This is a typical model leaderboard update, with the core takeaway being that Grok 4.5 currently holds the lead in this evaluation.

Related event: Grok 4.5 Tops Coding Q&A Benchmark with New Collaborative Release(3 posts)→

Original post →

More from Models

Models channel →