CursorBench 4.0: Grok 4.7 xhigh hits 46.3% at $6.01/task, big jump over 4.6 at same cost

haider1 · x · 2026-09-22

haider shares CursorBench 4.0 coding benchmark results: Grok 4.7 xhigh scores 46.3% at $6.01/task; Fable 5.1 medium 46.8% at $7.05; GPT-5.6 sol max 41.7% at $8.23. Compared to Grok 4.6 xhigh (41.4%), the new version jumps to 46.3% at essentially the same cost — a notable efficiency gain.

Original post →

More from Models

Models channel →