Grok 4.5 Tops Coding Benchmark with High Token Efficiency

xAI's Grok 4.5 paired with Grok Build scored 84 on the SWE-Atlas-QnA benchmark, tying for first with Codex GPT-5.6. The setup also demonstrated exceptional token efficiency, consuming significantly fewer tokens per task than mainstream competitors.

2026-07-10 ~ 2026-07-11 · 3 related posts

Full story(20 episodes)→