Untrained LFM2.5 logprob readout hits 0.710, trailing fine-tuned lev-350M's 0.725

helloiamleonie · x · 2026-09-24

After a developer ported jaredpalmer's kev to Liquid AI's LFM2.5-350M (dubbed lev), another experiment tested whether raw base-model probabilities alone could work: using llama.cpp 1-token logprob readout with no fine-tuning or decision head, LFM2.5-350M scored 0.638 and LFM2.5-2.6B scored 0.710, versus 0.725 for lev-350M. The gap between untrained models and lev is notable, showing how far raw probability readout already goes.

Original post →

More from Models

Models channel →