LFM2.5 8B A1B GGUF Models Model Generation Details This model was generated using llama.cpp at commit 94a220cd6 . Click here to get info on choosing the right GGUF model format Try LFM • Docs • LEAP • Discord LFM2.5 8B A1B LFM2.5 is a new family of hybrid models designed for on device deployment. It builds on the LFM2 architecture with extended pre training and reinforcement learning. On device personal assistant : Designed to power real life applications, chaining tool calls, and following complex instructions on all devices. Compressed performance : Competitive with much larger dense and MoE models on instruction following and agentic tasks. Unmatched throughput : Fastest in its size class on both CPU and GPU inference, with day one support for llama.cpp, MLX, vLLM, and SGLang. Find more information about LFM2.5 8B A1B in our blog post. AA Omniscience Index (higher is better) rewards correct answers and penalizes hallucinations. Scores range from 100 to 100. See more results on Artificial Analysis. 🗒️ Model Details Model Parameters Description LFM2.5 8B A1B Base 8.3B total / 1.5B active Pre trained base model for fine tuning LFM2.5 8B A1B 8.3B total / 1.5B active Reasoning tuned…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy