ik llama.cpp imatrix Quantizations of Qwen/Qwen3.6 27B NOTE ik llama.cpp can also run your existing GGUFs from bartowski, unsloth, mradermacher, etc if you want to try it out before downloading my quants. Only a couple quants in this collection are compatible with mainline llamma.cpp/LMStudio/KoboldCPP/etc as mentioned in the specific description, all others require ik llama.cpp. Some of ik's new quants are supported with Nexesenex/croco.cpp fork of KoboldCPP with Windows builds. Also check for ik llama.cpp windows builds by Thireus here.. These quants provide best in class perplexity for the given memory footprint. Big Thanks Shout out to Wendell and the Level1Techs crew, the community Forums, YouTube Channel! BIG thanks for providing BIG hardware expertise and access to run these experiments and make these great quants available to the community!!! Also thanks to all the folks in the quantizing and inferencing community on BeaverAI Club Discord and on r/LocalLLaMA for tips and tricks helping each other run, test, and benchmark all the fun new models! Thanks to huggingface for hosting all these big quants! Finally, I really appreciate the support from aifoundry.org so check out th…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy