Grok 4.7 falls to #24 on Vals Index, down 5 points from Grok 4.6
scaling01 · x · 2026-09-22
Vals AI's benchmark suite ranks Grok 4.7 #24 on the Vals Index at 54.2%, down 5.0 points from Grok 4.6 (#14, 59.2%), though still ahead of Grok 4.5 (#30, 51.5%). The biggest gains are in legal and medical work — but the overall regression right after release is notable.
More from Models
- Azure OpenAI content filter blocks 'S&M' — the standard finance shorthand for Sales & Marketing — peterjliu · 2026-09-22
- IFM's K2-Horizon-36B-A4B Matches 20x-Larger Models on AA Index Using New MoVA Architecture — victormustar · 2026-09-22
- Grok 4.7 fails again: $1.59 run produces laughable output — teortaxesTex · 2026-09-22
- LLM scam detection benchmarked: fitted TF-IDF baseline beats Jev, DeepSeek and local Qwen — justinbiebar · 2026-09-22
- Goodfire Finds DNA Model Evo 2 Encodes the Tree of Life as a Curved Activation Manifold — burny_tech · 2026-09-22
- Internal eval puts Grok 4.7 at #3 across 22 knowledge-work tasks for under $5 — realsohamparekh · 2026-09-22