xAI releases Grok 4.6, focused on long-running agents, matching GPT-5.6 Sol on AA index

thione · x · 2026-08-18

xAI officially released Grok 4.6, building on Grok 4.5 with a focus on long-running agents and more ambitious interactive and visual work — staying with complex multi-step tasks like research, codebase-wide work, or turning an idea into a polished application.

Benchmarks: Matches GPT-5.6 Sol on the Artificial Analysis Intelligence Index (composite of nine benchmarks); frontier-level on GDPVal-AA, DeepSWE 1.1, CursorBench 3.2 and FrontierCode 1.1.

Training: A longer supplemental training run than 4.5 with curated model-generated reasoning data, high-quality engineering data, and an improved optimizer/recipe; SFT trajectories were regenerated with Grok 4.5 across reasoning efforts, agent harnesses and domains, filtered via model-based checks.

Available today in Cursor and Grok Build with 2x included usage for the first week.

Original post →

More from coding & agent

coding & agent channel →