LFM2.5-2.6B released: full agent training pipeline compressed into 2.6B parameters

SergioPaniego · x · 2026-08-04

LFM2.5-2.6B model released, with training pipeline: base model → SFT → specialized teachers per domain (SFT + RLVR) → on-policy distillation back into one student → agentic RL.

Related event: LFM2.5-2.6B Released: Small Model Beats Larger Counterparts(3 posts)→

Original post →

More from Models

Models channel →