Negative Self-Distillation improves LLM reasoning by avoiding flawed reasoning paths

Rongcan Pei · hf · 2026-09-11

A new method, Negative Self-Distillation, improves LLM reasoning by pushing models away from self-generated flawed reasoning paths instead of imitating positive samples.

Original post →

More from Research

Research channel →