Gemma 3 12B IT Heretic v2 (Abliterated) 💬 Community: Join the Abliterlitics Discord for discussion, model releases and support. An abliterated version of Google's Gemma 3 12B IT created using Heretic v1.3.0. This model has reduced refusals while maintaining model quality, making it suitable as an uncensored text encoder for video generation models like LTX 2. Available in five quantization formats for ComfyUI — FP8, INT8 (ConvRot), NVFP4, and MXFP8 — covering everything from Ada GPUs to Blackwell. You can see the docker, scripts and configurations used to make these files on Heretic Docker Github. Update — July 11, 2026 Added INT8 (ConvRot) and MXFP8 quantization formats via convert to quant by silveroxides: INT8 ConvRot : Near lossless INT8 via SVD guided learned rounding (Prodigy optimizer), tensor wise scaling. Works on any GPU (Ampere+), natively supported in ComfyUI v0.27.0+. 13 GB. MXFP8 : Microscaling FP8 (OCP MX standard) with E8M0 per block power of 2 scales. Blackwell only. 13 GB. FP8 upgraded : Now uses per tensor scaling via convert to quant (previously naive unscaled cast). Pipeline now uses convert to quant for FP8, INT8, and MXFP8 quantization with ComfyUI compatibl…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy