[!NOTE] Includes Unsloth chat template fixes ! For llama.cpp , use jinja Unsloth Dynamic 2.0 achieves superior accuracy & outperforms other leading quants. Playground Playground Playground Leap LFM2 8B A1B LFM2 is a new generation of hybrid models developed by Liquid AI, specifically designed for edge AI and on device deployment. It sets a new standard in terms of quality, speed, and memory efficiency. We're releasing the weights of our first MoE based on LFM2, with 8.3B total parameters and 1.5B active parameters. LFM2 8B A1B is the best on device MoE in terms of both quality (comparable to 3 4B dense models) and speed (faster than Qwen3 1.7B). Code and knowledge capabilities are significantly improved compared to LFM2 2.6B. Quantized variants fit comfortably on high end phones, tablets, and laptops . Find more information about LFM2 8B A1B in our blog post. 📄 Model details Due to their small size, we recommend fine tuning LFM2 models on narrow use cases to maximize performance. They are particularly suited for agentic tasks, data extraction, RAG, creative writing, and multi turn conversations. However, we do not recommend using them for tasks that are knowledge intensive or requ…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy