shawnw3i/Huihui Qwen3.6 27B abliterated AWQ MTP This is an uncensored version of Qwen/Qwen3.6 27B created with abliteration (see remove refusals with transformers to know more about it). This is a crude, proof of concept implementation to remove refusals from an LLM model without using TransformerLens. Highlights AWQ Marlin kernel supported (auto converted by vLLM at runtime) MTP speculative decoding supported out of the box 110+ tok/s on a single A800 80GB (vLLM 0.21.0, MTP enabled, fp8 KV cache) vLLM Usage Warnings Risk of Sensitive or Controversial Outputs : This model’s safety filtering has been significantly reduced, potentially generating sensitive, controversial, or inappropriate content. Users should exercise caution and rigorously review generated outputs. Not Suitable for All Audiences : Due to limited content filtering, the model’s outputs may be inappropriate for public settings, underage users, or applications requiring high security. Legal and Ethical Responsibilities : Users must ensure their usage complies with local laws and ethical standards. Generated content may carry legal or ethical risks, and users are solely responsible for any consequences. Research and Exp…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy