Qwen3.5 35B A3B Revised GGUF Quantizations 2026 05 13 Update : Anyone who has downloaded the model prior to today should re download quants . Chat template overhaul, and uses my custom chat template. Revised GGUF quantizations of the official Qwen 3.5 35B A3B model, converted locally using llama.cpp b9009. These modified GGUFs have chat template fixes/improvements while maintaning original pure functionality. They're nicknamed Revised as it is a change to the original source. This model is no different capability wise from the Pure GGUF variant. Think of it as an improved version of it. This version comes with a baked in fixed chat template made by me, you can find it here. Revised Jinja2 chat templates for Qwen models. Drop in replacements for the official templates with bug fixes and quality of life improvements. — Smoffyy License & Attribution This model is based on Qwen 3.5 35B A3B, which was released under the Apache 2.0 license. This work and all derivatives are released under Apache 2.0. Should the original model's license change in the future, this version remains under Apache 2.0 in perpetuity. Why was this variant made? The official Qwen templates have bugs and missing fe…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy