HLE leaderboard: Grok 4.7 at #21 while Gemini 3.8 and Muse 1.3 lead by a margin
himanshustwts · x · 2026-09-22
According to a posted Humanity's Last Exam (HLE) leaderboard, Grok 4.7 currently sits at #21, while Gemini 3.8 and Muse 1.3 lead the chart by a clear margin. The model names and scores are third-party claims, not yet officially confirmed.
Related event: Grok 4.7 Ranks Only 21st on HLE Leaderboard(2 posts)→
More from Models
- Alignment backfires: model strips all faces from a deepfake detection dataset mid-task — generativist · 2026-09-22
- Grok 4.7 falls to #24 on Vals Index, down 5 points from Grok 4.6 — scaling01 · 2026-09-22
- Game Theory of Model Launch Dates: Launching Early Admits Your Model Is Weaker — cocktailpeanut · 2026-09-22
- Liquid AI's LFM2.5 tops mobile benchmarks: 2.32GB memory, 8s latency on iPhone 17 Pro — maximelabonne · 2026-09-22
- Jev reportedly does tensor logic under the hood: differentiable IF args, no wasted gen tokens — StewartalsopIII · 2026-09-22
- Multilingual Model Laya Trending on Hugging Face — convaiinnovations · 2026-09-22