Hugging Face Explores Async OPD for Faster Training

Hugging Face's post-training team discussed asynchronous on-policy distillation (AsyncOPD), highlighting its ability to accelerate training throughput by 2 to 3 times in math tasks.

2026-07-21 ~ 2026-07-21 · 2 related posts