Qwen3.6 35B A3B Uncensored HauhauCS MTP This model is a modified version of Qwen3.6 35B A3B Uncensored HauhauCS Aggressive grafted with the Multi Token Prediction (MTP) module using the MTP donor from the Qwen 3.6 35B A3B MTP GGUF series by Unsloth . This modification aims to provide faster inference speeds via MTP based speculative decoding without sacrificing the base model's original quality or capabilities. Specifications Base Model: HauhauCS/Qwen3.6 35B A3B Uncensored HauhauCS Aggressive MTP Donor: unsloth/Qwen3.6 35B A3B MTP GGUF Architecture: Mixture of Experts (MoE) — 35B total parameters / ~3B active per forward pass (256 experts, 8 routed per token) Context Window: 262K (262,144 tokens) Multimodal Capabilities: Supports text, image, and video processing Uncensored Nature: Inherits the Aggressive variant from HauhauCS (0/465 refusals on standard evaluation datasets, removing default refusal behavior while maintaining base performance and model traits). Key MTP Features Inference Speedup: Offers an estimated speedup of 1.4x to 2.2x faster generation (depending on hardware specifications and the inference backend). Consistent Quality: Retains the same output distribution as…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy