SVD initialization of LoRA accelerates RL training convergence

Trajectory Labs shows that initializing LoRA adapters with top-k SVD components of pretrained weights (PiSSA), instead of the community's common approach, speeds up RL training convergence and improves downstream performance.

2026-09-17 ~ 2026-09-17 · 2 related posts