Qwen3 VL 4B Instruct Heretic (GGUF) 💬 Community: Join the Abliterlitics Discord for discussion, model releases and support. GGUF quantisations of the Heretic abliteration of Qwen3 VL 4B Instruct. The full forensic report comparing all candidate variants sits in the main repo. Available formats This model is published across three repos. Pick the one that matches your runtime. Repo Best for Contents Qwen3 VL 4b Heretic transformers, vLLM, HF Hub bf16 weights with config, vision encoder preserved Qwen3 VL 4b Heretic GGUF (this repo) llama.cpp, Ollama, LM Studio, ComfyUI GGUF GGUF quants from Q3 K M up to F16 (text path) Qwen3 VL 4b Heretic ComfyUI ComfyUI text encoder bf16, fp8, int8, nvfp4 and mxfp8 checkpoints Why this variant? Several Heretic trials were run against Qwen3 VL 4B and all of them reach 100% HarmBench ASR, up from 30.8% on the base. This build was picked because, with safety tied, it wins on the tie breakers: Base Heretic HarmBench ASR 30.8% 100% KL divergence (lower is better) 0.0283 (lowest of the candidates) GSM8K 78.62% 77.18% ( −1.83% , smallest drop) MMLU 69.58% 69.61% (+0.03%) Tensors changed 54 (pure rank 1) See the full report for the comparison. Files File…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy