NAVER's DLoop Speeds Up Speculative Decoding by up to 41%
NAVER AI Lab's DLoop lets a draft model loop multiple drafting rounds before verification, cutting unnecessary target-model forward passes and delivering 5-41% faster inference.
2026-10-08 ~ 2026-10-09 · 2 related posts