Grok 4.7 launches at $2/$6 per M tokens, hits 46.3% on CursorBench 4.0
Scobleizer · x · 2026-09-22
SpaceXAI releases Grok 4.7, its most capable model for coding and knowledge work.
- Larger base model, longer RL on multi-hour tasks; better self-verification and long-context handling
- Same price as 4.6: $2/$6 per million tokens — a fraction of GPT-5.6 Sol ($4/$20) and Fable 5.1 ($10/$50)
- Benchmarks: CursorBench 4.0 xhigh 46.3% (vs 40.4% for 4.6), Terminal-Bench 4.0 38.0% (nearly double), Harvey Legal 19.6%, EEBench 64.0%
- Fable 5.1 still leads some software benches, but 4.7 closes much of the gap far cheaper
Live in Cursor, Grok Build, and the API, with a new safeguard stack for refusals and dual-use safety.
Related event: xAI Launches Grok 4.7, Its Strongest Coding Model Yet(26 posts)→
More from Models
- Azure OpenAI content filter blocks 'S&M' — the standard finance shorthand for Sales & Marketing — peterjliu · 2026-09-22
- IFM's K2-Horizon-36B-A4B Matches 20x-Larger Models on AA Index Using New MoVA Architecture — victormustar · 2026-09-22
- Grok 4.7 fails again: $1.59 run produces laughable output — teortaxesTex · 2026-09-22
- LLM scam detection benchmarked: fitted TF-IDF baseline beats Jev, DeepSeek and local Qwen — justinbiebar · 2026-09-22
- Goodfire Finds DNA Model Evo 2 Encodes the Tree of Life as a Curved Activation Manifold — burny_tech · 2026-09-22
- Internal eval puts Grok 4.7 at #3 across 22 knowledge-work tasks for under $5 — realsohamparekh · 2026-09-22