Grok 4.7 xHigh hits 58% on AA-Briefcase, just 1 point behind Claude Fable 5.1 Max
XFreeze · x · 2026-09-22
Grok 4.7 xHigh now sits at the top of Artificial Analysis' AA-Briefcase benchmark with 58%, just one point behind Claude Fable 5.1 Max at 59%. Grok outscores GPT-6 Astra, GPT-5.6, Gemini, Kimi, GLM and nearly every other frontier model on the benchmark, marking the closest race yet at the top.
More from Models
- Xiaomi open-sources MiMo-V2.6 omni-modal models, topping open-model index at 46.32 — victormustar · 2026-09-22
- SemiAnalysis says open source is dying, yet 20+ open models shipped in the past month — _lewtun · 2026-09-22
- Grok 4.7 posts 59% recall on defensive cyber bench at half the cost of rivals — andreamichi · 2026-09-22
- Grok 4.7 jumps from #9 to #3 on BuildingBench with 0.783, 66% cheaper than Fable 5.1 — ZhitingHu · 2026-09-22
- Jev explained: why the AI community's new favorite isn't a traditional LLM — multiply_matrix · 2026-09-22
- LLMs excel at 1-token output — dev proposes replacing low/medium/high reasoning tiers with token counts — arkuto · 2026-09-22