User tests Grok 4.7 and calls it worse than the Gemini series
prvthvm · reddit · 2026-09-22
A Reddit user reports being deeply disappointed after trying Grok 4.7, saying they never expected a model worse than the Gemini series. They note Elon Musk had previously tweeted the model would beat Astra and 5.1 across all benchmarks. Purely a personal take with no detailed test data.
More from Models
- Vals AI: Grok 4.7 drops to #24 on Vals Index, down 5 points from Grok 4.6 — zacharynado · 2026-09-22
- Grok 4.7 example: three-year financial analysis exposes currency-masked growth stall — ArtificialAnlys · 2026-09-22
- AA example: Grok 4.7 independently runs valuation chain and flags divergence from deal partner — ArtificialAnlys · 2026-09-22
- Grok 4.7 ranks just behind Anthropic's Opus 5 on AA-Briefcase at ~50% of the cost per task — ArtificialAnlys · 2026-09-22
- Replicating ExploitBench Would Cost ~$59.3M in API Fees, Security Researcher Estimates — OwariDa · 2026-09-22
- Xiaomi Releases Small Qwen 3.5 9B Distill SFT'd on MiMo Data, Plus RL Environments — teortaxesTex · 2026-09-22