Salesforce: Imitating Stronger Models' Trajectories Hurts Agent Performance

Salesforce AI's new paper finds that fine-tuning a weak agent on a stronger model's full trajectories degrades performance by 4-30 points, while online error-correction-style fine-tuning is the effective approach.

2026-09-22 ~ 2026-09-22 · 2 related posts