Try LFM • Docs • LEAP • Discord LFM2.5 8B A1B LFM2.5 is a new family of hybrid models designed for on device deployment. It builds on the LFM2 architecture with extended pre training and reinforcement learning. On device personal assistant : Designed to power real life applications, chaining tool calls, and following complex instructions on all devices. Compressed performance : Competitive with much larger dense and MoE models on instruction following and agentic tasks. Unmatched throughput : Fastest in its size class on both CPU and GPU inference, with day one support for llama.cpp, MLX, vLLM, and SGLang. Find more information about LFM2.5 8B A1B in our blog post. AA Omniscience Index (higher is better) rewards correct answers and penalizes hallucinations. Scores range from 100 to 100. See more results on Artificial Analysis. 🗒️ Model Details Model Parameters Description LFM2.5 8B A1B Base 8.3B total / 1.5B active Pre trained base model for fine tuning LFM2.5 8B A1B 8.3B total / 1.5B active Reasoning tuned general purpose model LFM2.5 8B A1B is a general purpose text only model with the following features: Total parameters : 8.3B Active parameters : 1.5B Number of layers : 24 (18…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy