DeepSeek V4 Pro benchmarks close to Opus 5 on KernelBench-Hard

teortaxesTex · x · 2026-08-21

Benchmark results show DeepSeek V4 Pro achieving 9.35% of roofline performance on the TopK task in KernelBench-Hard, closely trailing Opus 5 at 9.46%. The data also compares performance across models like Kimi K3, Fable 5, and Qwen 3.8 Max on various tasks including FP8, KDA, and Paged attention.

Original post →

More from Models

Models channel →