Try LFM • Docs • LEAP • Discord LFM2‑VL LFM2 VL 3B is the newest and most capable model in Liquid AI's multimodal LFM2 VL series, designed to process text and images with variable resolutions. Built on the LFM2 backbone, it extends the architecture for higher capacity reasoning and stronger visual understanding while retaining efficiency. We are releasing the weights of the new 3B checkpoint—offering higher performance across benchmarks while remaining optimized for scalable deployment. Competitive multimodal performance among lightweight open models. Enhanced visual understanding and reasoning , particularly on fine grained perception tasks Retains efficient inference with the same flexible architecture and user tunable speed quality tradeoffs Processes native resolutions up to 512×512 with intelligent patch based handling for larger inputs For more details, see the LFM2 VL 3B post and the LFM2 blog post. 📄 Model details Due to their small size, we recommend fine tuning LFM2 VL models on narrow use cases to maximize performance. They were trained for instruction following and lightweight agentic flows. Not intended for safety‑critical decisions. Property LFM2 VL 450M LFM2 VL 1.6B…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy