Grok 4.5 Ties with GPT-5.6 in Benchmark

scaling01 · x · 2026-07-10

The post shares a KernelBench-Hard update, noting that Grok 4.5 high and GPT-5.6 Sol xhigh achieved roughly tied overall scores on the RTX PRO 6000.

It further breaks down performance across multiple sub-categories, including FP8, paged attention, W4A16, and KDA, highlighting the strengths and weaknesses of each model across different operators and Triton/cuBLAS pathways.

Original post →

More from Models

Models channel →