Shanghai AI Lab Proposes SCALE: Entropy-Gated Control to Reverse SFT Features

Shanghai-AI-Laboratory · hf · 2026-10-05

Shanghai AI Lab's paper shows existing token-reweighting methods in SFT can only suppress or amplify updates, never reverse harmful learned features. Their SCALE method freezes the pretrained model and SFT delta, learning bounded token/module gates via predictive entropy alone to suppress, reverse, or extrapolate SFT features — beating baselines on Qwen math models (37.84/43.60/36.57) while retaining general and code performance.

Original post →

More from Research

Research channel →