Llamacpp imatrix Quantizations of Qwen3 14B abliterated by huihui ai Using llama.cpp release b5284 for quantization. Original model: https://huggingface.co/huihui ai/Qwen3 14B abliterated All quants made using imatrix option with dataset from here Run them in LM Studio Run them directly with llama.cpp, or any other llama.cpp based project Prompt format No chat template specified so default is used. This may be incorrect, check original model card for details. Download a file (not the whole branch) from below: Filename Quant type File Size Split Description Qwen3 14B abliterated bf16.gguf bf16 29.54GB false Full BF16 weights. Qwen3 14B abliterated Q8 0.gguf Q8 0 15.70GB false Extremely high quality, generally unneeded but max available quant. Qwen3 14B abliterated Q6 K L.gguf Q6 K L 12.50GB false Uses Q8 0 for embed and output weights. Very high quality, near perfect, recommended . Qwen3 14B abliterated Q6 K.gguf Q6 K 12.12GB false Very high quality, near perfect, recommended . Qwen3 14B abliterated Q5 K L.gguf Q5 K L 10.99GB false Uses Q8 0 for embed and output weights. High quality, recommended . Qwen3 14B abliterated Q5 K M.gguf Q5 K M 10.51GB false High quality, recommended . Qw…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy