Qwen3.6 35B Abliterated Heretic — 4 bit AWQ Quantized Multimodal MoE A 4 bit AWQ quantized multimodal Mixture of Experts (MoE) model derived from Youssofal/Qwen3.6 35B A3B Abliterated Heretic BF16 . This is a quantized variant of Youssofal's Abliterated Heretic Uncensored model — an unrestricted, uncensored iteration of the Qwen3.6 35B family — compressed to 4 bit via Activation aware Weight Quantization for deployment friendly inference. Model Type & Architecture Property Value Architecture Qwen3 5MoeForConditionalGeneration Model Family Qwen3.6 35B Abliterated Heretic Uncensored (quantized) Total Parameters ~36.2B Text Hidden Size 2,048 Vision Hidden Size 1,152 Text Layers 40 (hybrid: 3 linear + 1 full attention repeating) Attention Grouped Query Attention (GQA), 16 query heads, 2 KV heads (ratio 8:1) Head Dimension 256 Vocabulary Size 248,320 Max Context Length 262,144 tokens MoE Experts 256 per layer, 8 activated per token MoE Intermediate Size 512 RoPE Theta 10,000,000 (long context optimized) RoPE Interleaved multi RoPE sections [11, 11, 10] for multimodal Activation SiLU (text), GELU Tanh (vision) Transformers Version 5.6.0.dev0 Vision Encoder Property Value Layers 27 Hidden…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy