Sharpening Tax: Post-Training Sharpens Old Skills Rather Than Teaching New Ones

Researchers including Meta propose the "Sharpening Tax," showing that RL post-training mainly sharpens existing skills: it boosts pass@1 but sacrifices pass@K coverage compared to pretrained models with a lightweight inference harness.

2026-10-02 ~ 2026-10-02 · 2 related posts