Try LFM • Docs • LEAP • Discord LFM2.5 1.2B Thinking LFM2.5 is a new family of hybrid models designed for on device deployment . It builds on the LFM2 architecture with extended pre training and reinforcement learning. Best in class performance : A 1.2B model rivaling much larger models, bringing high quality AI to your pocket. Fast edge inference : 239 tok/s decode on AMD CPU, 82 tok/s on mobile NPU. Runs under 1GB of memory with day one support for llama.cpp, MLX, and vLLM. Scaled training : Extended pre training from 10T to 28T tokens and large scale multi stage reinforcement learning. Find more information about LFM2.5 1.2B Thinking in our blog post. 🗒️ Model Details Model Parameters Description LFM2.5 1.2B Base 1.2B Pre trained base model for fine tuning LFM2.5 1.2B Instruct 1.2B General purpose instruction tuned model LFM2.5 1.2B Thinking 1.2B General purpose reasoning model LFM2.5 1.2B JP 1.2B Japanese optimized chat model LFM2.5 VL 1.6B 1.6B Vision language model with fast inference LFM2.5 Audio 1.5B 1.5B Audio language model for speech and text I/O LFM2.5 1.2B Thinking is a general purpose text only model with the following features: Number of parameters : 1.17B Number of…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy