"700 Zone": This model's performance has been "beaten" (by a mile) by the new "711" (and includes "MTP" quants too): https://huggingface.co/DavidAU/Qwen3.6 27B Fable Fusion 711 Uncensored Heretic NM DAU NEO MAX MTP GGUF This model's benchmarks are in the OpenAI, Claude and Gemini "zone" it is that strong. All Optimized Quants BENCHMARKED (5 metrics): You know exactly how strong each NEO CODE quant is relative to full precision model. IQ4XS stands out at 94% of full precision at 1/4 the size of the full model. Then there are the Q6 and Q8, with the latter hitting 98.38% (of full precision). Even the lowest (IQ2 M) at 1/5 (10.5GB) the size of the full model scores 82.66%! All quants and metrics listed below. Uncensored quants are here (exceeds all metrics of the quants at this repo too): Click here Qwen3.6 27B NEO CODE Di IMatrix MAX GGUF Team Qwen exceeded all expectations with this 27B model [even exceeding their own 398B model] AND GEMMA 4s too, so here are the quants to match. All GGUFs benchmarked against full precision model (also primer below) and there is an Unsloth VS DavidAU GGUF showdown too. (You guys know I love the Unsloth team.) Quant "engineering" focused on balance a…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy