Self-OPD: On-Policy Distillation for Flow Matching without Teachers

Shiyi Zhang · hf · 2026-08-28

Self-OPD is a new method for flow matching models designed to eliminate task-specific teachers. It optimizes the velocity field for multi-objective alignment using self-explored stochastic branches and normalized advantages.

Original post →

More from Research

Research channel →