Qwen3.6 27B Abliterated + MTP GGUF The first publicly available Qwen3.6 27B uncensored GGUF with native MTP speculative decoding. Refusal free at the weight level · Full MTP block grafted · ~70 t/s on RTX 3090 · No custom fork required Published by Gastón Parravicini Why this exists When Qwen3.6 27B dropped, two things were true at the same time: The only abliterated versions available had no MTP support — the draft heads were stripped during the merge, killing speculative decoding speed The only MTP enabled GGUFs were fully censored — original refusal behavior intact Nobody had combined both. Doing it required writing custom patches to handle Qwen3.6's MTP tensor naming conventions and avoid GGUF metadata corruption. This release is the result of that work. This was the first. It still has the most complete quant coverage. What this release adds Refusal suppression Removed at the weight level using two pass orthogonal projection abliteration (abliterix + Optuna TPE). KL divergence of 0.024 vs the base model — well below the 0.05 threshold where quality degradation becomes measurable. General intelligence, reasoning, and tool use are fully intact. Full MTP speculative decoding The…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy