Internal eval puts Grok 4.7 at #3 across 22 knowledge-work tasks for under $5
realsohamparekh · x · 2026-09-22
A third-party evaluated Grok 4.7 on an internal sample of 22 knowledge-work tasks spanning engineering, bizops, finance and healthcare: it ranked #3 on the table at a cost under $5.
- Strong at multi-document policy reasoning and precedence handling
- Good at large structured audits across CSV, JSON, Markdown and spreadsheets
- Consistently produced the requested deliverable files
The release notes Grok 4.7 is a notable improvement over Grok 4.6 at the same price and speed.
Related event: Grok-4.7 Ranks Third Across 22 Knowledge Tasks, Costs Under $5(2 posts)→
More from Models
- Paradigm launches Limite 1B - Violetto, a model for high-frequency mathematical intelligence — tensorqt · 2026-09-22
- Reliquary-4B: A 4B math & code model trained via decentralized RL with community rollouts — const_reborn · 2026-09-22
- Users say they can't trick Jev into hallucinating — BLUECOW009 · 2026-09-22
- Measured trade-offs of three REAP-pruned Qwen3.8-Flash-Next MLX builds on Apple Silicon — MensaProdigy · 2026-09-22
- Dev claims further-optimized DeepSeek V4 NVFP4 uses 190GB of 192GB VRAM — HankYeomans · 2026-09-22
- OpenAI researcher Will Depue on why voice models still lack true realtime chat — willdepue · 2026-09-22