Grok 4.7 out: Terminal-Bench jumps 20.3% to 38%, beats GPT-5.6 Sol on CursorBench
thesaraharminta · x · 2026-09-22
- Grok 4.7 is out, with its biggest gain on Terminal-Bench: 20.3% → 38%.
- It beats GPT-5.6 Sol on CursorBench (46.3% vs 41.7%), while token pricing stays the same as Grok 4.6.
- It still trails Fable 5.1 on both benchmarks — a solid but non-leading jump.
Related event: xAI Launches Grok 4.7, Its Strongest Coding Model Yet(26 posts)→
More from Models
- Grok 4.7 fails again: $1.59 run produces laughable output — teortaxesTex · 2026-09-22
- LLM scam detection benchmarked: fitted TF-IDF baseline beats Jev, DeepSeek and local Qwen — justinbiebar · 2026-09-22
- Goodfire Finds DNA Model Evo 2 Encodes the Tree of Life as a Curved Activation Manifold — burny_tech · 2026-09-22
- Internal eval puts Grok 4.7 at #3 across 22 knowledge-work tasks for under $5 — realsohamparekh · 2026-09-22
- Codex code leak hints at GPT-6 Luna with pricing already in place, alongside GPT-6 Sol — haider1 · 2026-09-22
- Alex Atallah: specialized model variants could spark a 'Jev moment' and break provider lock-in — multiply_matrix · 2026-09-22