RL Post-Training Updates Can Be Sparsely Decomposed

menhguin · x · 2026-07-15

This repost covers new research on RL post-training, addressing the core issue: while RL updates are effective, the parameter changes themselves act as a "black box."

The paper proposes treating the genuinely effective "reasoning component" of RL as a compact reconnection matrix within the base model's spectral space. Based on this, they introduce SAR: a retraining-free post-processing method that projects raw RL updates onto this "reasoning core" to better understand, purify, and merge RL-trained models.

Key conclusions from the text include:

Original post →

More from Research

Research channel →