Qwen3.6 35B A3B DFlash GGUF quantizations of z lab/Qwen3.6 35B A3B DFlash. Converted to BF16 using convert hf to gguf.py , then quantized using llama quantize from llama.cpp. Available quants Quant Bits Size Notes Q4 K M 4 ~235 MB Average quality Q5 K 5 ~280 MB High quality Q6 K 6 ~326 MB Very high quality Q8 0 8 ~421 MB Highest quality, near lossless, Recommended BF16 16 ~771 MB Full precision, reference file Usage Use in conjunction with existing Qwen3.6 Quants, example config if using llama server : Original model See the original model card for details on capabilities, benchmarks, and license.
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy