Qwen2.5 7B Instruct Abliterated (GGUF) An abliterated (uncensored) version of Qwen/Qwen2.5 7B Instruct in GGUF format, ready for local inference with llama.cpp, Ollama, or LM Studio. Abliteration removes the refusal training from the model while preserving its core capabilities — useful for research, creative writing, and scenarios where you need unrestricted model output. Quick Start With Ollama With llama.cpp With LM Studio Search for richardyoung/Qwen2.5 7B Instruct abliterated GGUF in the model browser, or download manually and import. Available Quantizations Quantization Use Case Q2 K Minimum RAM, lower quality Q4 K M Recommended — good balance of quality and speed Q5 K M Higher quality, more RAM Q6 K Near original quality Q8 0 Maximum quality, most RAM What is Abliteration? Abliteration is a technique that identifies and removes the "refusal direction" in a model's residual stream. Unlike fine tuning, it surgically modifies the model's behavior without retraining, preserving the original model's knowledge and capabilities. For more details, see the original research: Refusal in Language Models Is Mediated by a Single Direction Intended Use This model is intended for: Research…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy