Multi-domain on-policy distillation lets one model learn many skills at once

helloiamleonie · x · 2026-09-25

Leonie explains MOPD (multi-domain on-policy distillation) used in LFM2.5 training: extending on-policy distillation to multiple teachers so the student model learns different skills simultaneously, rather than sequentially. It's the key step that merges specialized teachers' abilities back into a single checkpoint.

Related event: Liquid AI Releases LFM2.5-2.6B and Details Its On-Device Training Recipe(10 posts)→

Original post →

More from Research

Research channel →