LiquidAI Launches LFM2.5-VL-3B: A Lightweight Vision-Language Model for Screen Understanding

helloiamleonie · x · 2026-08-12

LiquidAI has released LFM2.5-VL-3B, a lightweight vision-language model designed to read digital screens, documents, and the physical world.

Built on the LFM2.5-2.6B base with a SigLIP2 vision encoder, the model excels in screen understanding, OCR, and tool use across mobile, web, and desktop. It delivers performance comparable to or better than models up to 2.6x its size, achieving high scores on benchmarks like ScreenSpot-v2 and RealWorldQA.

Related event: LiquidAI Releases LFM2.5-VL-3B Edge Multimodal Model(7 posts)→

Original post →

More from Models

Models channel →