Try LFM • Docs • LEAP • Discord LFM2.5‑VL 450M LFM2.5‑VL 450M is Liquid AI's refreshed version of the first vision language model, LFM2 VL 450M, built on an updated backbone LFM2.5 350M and tuned for stronger real world performance. Find more about LFM2.5 family of models in our blog post. Enhanced instruction following on vision and language tasks. Improved multilingual vision understanding in Arabic, Chinese, French, German, Japanese, Korean, Portuguese and Spanish. Bounding box prediction and object detection for grounded visual understanding. Function calling support for text only input. 🎥⚡️ You can try LFM2.5 VL 450M running locally in your browser with our real time video stream captioning WebGPU demo 🎥⚡️ Alternatively, try the API model on the Playground. 📄 Model details LFM2.5 VL 450M is a general purpose vision language model with the following features: LM Backbone : LFM2.5 350M Vision encoder : SigLIP2 NaFlex shape‑optimized 86M Context length : 32,768 tokens Vocabulary size : 65,536 Languages : English, Arabic, Chinese, French, German, Japanese, Korean, Portuguese, and Spanish Native resolution processing : handles images up to 512 512 pixels without upscaling and pr…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy