Smaller Models Make Better Rejects: Study Rethinks Preference Distillation from 7B to 72B

LinkedIn · hf · 2026-10-02

LinkedIn research challenges two default assumptions in preference distillation: that self-generated failures are the most informative negatives, and that rejects must come from models at least as large as the student.

Original post →

More from Research

Research channel →