HPD-Parsing reaches 94.91% on OmniDocBench v1.6 with 4,752 TPS throughput

pmttyji · reddit · 2026-07-23

HPD-Parsing proposes a hierarchical parallel decoding approach for document parsing.

The model splits work between a global layout branch and localized content branches, then further reduces decoding with Progressive Multi-Token Prediction and shared-prefix KV cache reuse. According to the post, the 1B model reaches 94.91% on OmniDocBench v1.6 and peaks at 4,752 TPS, which the authors say is 2.62× faster than the previous best parser and 3.06× faster than its autoregressive baseline.

Related event: HPD-Parsing Achieves High-Throughput Document Parsing via Hierarchical Parallel Decoding(2 posts)→

Original post →

More from Research

Research channel →