Qwen3.6 35B A3B — Abliterated (GGUF) GGUF format of wangzhang/Qwen3.6 35B A3B abliterated for use with llama.cpp, Ollama, LM Studio, and other GGUF compatible tools. Available Formats File Format Size Notes Qwen3.6 35B A3B abliterated BF16.gguf BF16 65 GB Full precision, best quality Qwen3.6 35B A3B abliterated Q8 0.gguf Q8 0 35 GB Near lossless, fits 48GB GPU Qwen3.6 35B A3B abliterated Q4 K M.gguf Q4 K M 20 GB Good balance, fits 24GB GPU Performance Metric Value Refusals (LLM judge, 100 eval prompts) 7/100 KL divergence from base 0.0189 Baseline refusals (original model) 100/100 LLM judge model google/gemini 3 flash preview See the full model card for detailed methodology, evaluation standards, and usage instructions. Usage with llama.cpp Usage with Ollama Disclaimer This model is released for research purposes only. The abliteration process removes safety guardrails — use responsibly.
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy