DepthBench aims to settle the race to beat the 'depth curse' in deep LLMs

FinanceYF5 · x · 2026-09-30

Deeper isn't automatically better: Kimi K3 uses AttnRes, DeepSeek V4 uses mHC, and ByteDance proposed HC — all attacking the 'depth curse'. Since these methods differ in training budget, model design and codebase, they can't be compared directly. The author introduces DepthBench, a benchmark to systematically evaluate which depth-scaling approach actually works.

Original post →

More from Models

Models channel →