xAI launches Grok 4.7: twice as fast at half the price, tops coding benchmarks
tetsuoai · x · 2026-09-22
- xAI released Grok 4.7, its most capable model for coding and knowledge work, built on a larger new base model with a longer RL run focused on multi-hour tasks, better self-verification, and longer context handling.
- Priced identically to Grok 4.6 at $2/M input and $6/M output tokens — half the price of GPT-5.6 Sol ($4/$20) and Fable 5.1 Max ($10/$50), claimed twice as fast.
- Benchmarks: 46.3% on CursorBench 4.0 (vs 40.4% for 4.6), 64.0% on EEBench, 38.0% on Terminal-Bench 4.0 (vs 20.3%), 19.6% on Harvey Legal Agent Benchmark where GPT-5.6 Sol scores just 2.5%, and 56.7% on HealthBench Professional.
- Ships with the company's best-calibrated safeguards to date.
Related event: xAI launches Grok 4.7: same price, faster speed, big benchmark jumps(33 posts)→
More from Models
- Gemini beats GPT-6 Astra at robot capture the flag, winning 70% of matches — chris_j_paxton · 2026-09-22
- 105 Planted Bugs Put Grok 4.7 at 28.7 vs GPT-6 Astra's 45 in Real-Repo Coding Test — PawelHuryn · 2026-09-22
- Regression on OpenAI's GPT-5.6 Benchmark Data Backs Out Their λ Values — tobyordoxford · 2026-09-22
- Developer Burns 1.46B Tokens in 13 Hours, Jokes He's "Part of the Infrastructure" — MaziyarPanahi · 2026-09-22
- Swarm scaling needs squared inference to match chain-of-thought gains, analysis of OpenAI curves finds — tobyordoxford · 2026-09-22
- User claims Gemini 4 Pro in production isn't the model on benchmarks — Ambroverse · 2026-09-22