Grok 4.6 review: capability up but latency doubles, experience worse than 4.5
brandon_galang · x · 2026-08-15
brandongalang quotes stanine's evaluation showing Grok 4.6 pass rate dropped from 87.3% to 85.9% on RipplingBench, median latency nearly doubled from 71s to 131s, with more errors and refusals. He personally finds 4.6 better at code but overall feel worse than 4.5, looking forward to 4.7.
More from Models
- Tencent Releases UI-Mate-27B, a Desktop GUI Agent Model — tencent · 2026-08-24
- Sakana AI translation outperforms Google and DeepL in Japanese-English benchmarks — SakanaAILabs · 2026-08-24
- Developer haider makes his own LLM tier list after disagreeing with theo's rankings — haider1 · 2026-08-24
- Mystery OxAlpha Beats Claude; Alibaba Raises $10B for AI — 创业邦 · 2026-08-24
- OpenAI and Google cut LLM prices; mystery OxAlpha model beats Claude on DeepSWE — 创业邦 · 2026-08-24
- AI News Digest: DeepSeek Weekend Discounts, GPT-5.6 Sol Price Cut, Alibaba's $10B AI Raise — APPSO · 2026-08-24