Liquid AI Launches LFM2.5-VL-3B: A Lightweight Vision-Language Model Outperforming 2.6x Larger Rivals

JosephJacks_ · x · 2026-08-13

Liquid AI has released LFM2.5-VL-3B, a lightweight vision-language model. Built on the LFM2.5-2.6B base and a SigLIP2 vision encoder, it was pre-trained on 34T tokens.

The model specializes in reading screens, documents, and the physical world across mobile, web, and desktop platforms. It supports tool calling from both text and image inputs. In benchmarks, it delivers impressive results:

Achieving performance comparable to or better than models up to 2.6x its size, it serves as a highly efficient base for custom applications.

Related event: LiquidAI Releases LFM2.5-VL-3B Edge Multimodal Model(7 posts)→

Original post →

More from Models

Models channel →