Qwen3.6 12B IQ Ultra Heretic Uncensored Thinking V2 Hightop - GGUF
GGUF quantizations of DavidAU/Qwen3.6-12B-IQ-Ultra-Heretic-Uncensored-Thinking-V2-Hightop.
Converted and quantized using llama.cpp b9192.
Available Quants
| Quant | Size | Quality |
|---|---|---|
| Q8_0 | 11.57 GB | Near-perfect |
| Q6_K | 8.93 GB | Excellent |
| Q5_K_M | 7.84 GB | Very good |
| Q5_K_S | 7.65 GB | Very good |
| Q5_0 | 7.65 GB | Good |
| Q4_K_M | 6.82 GB | Best balance |
| IQ4_NL | 6.58 GB | Very good (IQ) |
| Q4_K_S | 6.49 GB | Good |
| Q4_0 | 6.43 GB | Good |
| IQ4_XS | 6.31 GB | Good (IQ) |
| Q3_K_L | 5.94 GB | Acceptable |
| Q3_K_M | 5.58 GB | Acceptable |
| IQ3_M | 5.33 GB | Acceptable (IQ) |
| IQ3_S | 5.27 GB | Acceptable (IQ) |
| Q3_K_S | 5.15 GB | Fair |
| Q2_K | 4.60 GB | Minimal usable |
Original Model
DavidAU/Qwen3.6-12B-IQ-Ultra-Heretic-Uncensored-Thinking-V2-Hightop
Usage
Use with LM Studio, llama.cpp, Ollama, or any GGUF-compatible inference engine.