Grok 4.5 tops a new HighWalk benchmark on Laravel commit updates
elonmusk · x · 2026-07-29
- A repost claims Grok 4.5 took the top spot on a new benchmark called HighWalk.
- The benchmark reportedly tests whether models can update real technical specs from 46 Laravel commits, stressing code analysis, abstraction, and precise writing.
- In the shared results, Grok 4.5 (high) ranks #1 on overall quality-efficiency balance, Claude Opus 5 (high) has the highest raw quality with no hard failures, and GLM 5.2 is highlighted as the strongest open-weight model.
- The chart also says higher reasoning effort did not always help.
Related event: Grok 4.5 Tops HighWalk Benchmark(2 posts)→
More from Models
- Users say Laguna s2.1 still loops and misses tool calls after a strong launch — Possible_Grocery8079 · 2026-07-29
- GPT-5.6 Pro impresses as a code reviewer and bug hunter, says one user — dejavucoder · 2026-07-29
- From GPT-2 to KimiK3, a thread argues the story is bigger than scale — algo_diver · 2026-07-29
- Leaked video says Gemini 4 may be nearing release, showing physics and animation demos — WorldofAI · 2026-07-29
- Users say Anthropic’s Opus 5 has become nearly unreadable after personalization changes — himanshustwts · 2026-07-29
- Scobleizer says Grok 4.5 is the best coding model right now — Scobleizer · 2026-07-29