Grok 4.7 lands in Grok Build: 4 reasoning tiers, Terminal-Bench jumps 20% to 38%
XFreeze · x · 2026-09-22
Grok 4.7 is live in Grok Build with selectable Low/Medium/High/Extra High reasoning effort. Official benchmarks vs Grok 4.6: CursorBench 4.0 40.4%→46.3%, Terminal-Bench 4.0 20.3%→38.0%, EEBench 53.0%→64.0%, HealthBench Professional 48.5%→56.7%. Same price and speed, with a larger base model, longer RL training, better self-verification, and much stronger long-horizon reasoning.
Related event: xAI Launches Grok 4.7, Its Strongest Coding Model Yet(26 posts)→
More from Models
- Azure OpenAI content filter blocks 'S&M' — the standard finance shorthand for Sales & Marketing — peterjliu · 2026-09-22
- IFM's K2-Horizon-36B-A4B Matches 20x-Larger Models on AA Index Using New MoVA Architecture — victormustar · 2026-09-22
- Grok 4.7 fails again: $1.59 run produces laughable output — teortaxesTex · 2026-09-22
- LLM scam detection benchmarked: fitted TF-IDF baseline beats Jev, DeepSeek and local Qwen — justinbiebar · 2026-09-22
- Goodfire Finds DNA Model Evo 2 Encodes the Tree of Life as a Curved Activation Manifold — burny_tech · 2026-09-22
- Internal eval puts Grok 4.7 at #3 across 22 knowledge-work tasks for under $5 — realsohamparekh · 2026-09-22