Qwopus3.5 122B A10B Kimi K2.6 Distill Healed Abliterated Custom GGUF Quantizations CRITICAL COMPATIBILITY WARNING These are iqk format quantizations and are EXCLUSIVE to the ik llama.cpp fork. They will NOT work on mainline llama.cpp , standard LM Studio, standard Text Generation WebUI, or KoboldCPP. You must compile and run this using ikawrakow's llama.cpp fork, or a UI where you have manually swapped the backend to an ik llama.cpp build. This repository contains custom, mixed precision ik llama.cpp GGUF quantizations for OpenYourMind/Qwopus3.5 122B A10B Kimi K2.6 destill healed abliterated, a Kimi K2.6 distilled, healed, abliterated Qwen3.5 122B A10B MoE model. These quants use different precision levels for different layer types, keeping attention, SSM, shared expert, output, and MTP/NextN tensors at higher precision while compressing the routed experts, which make up the bulk of the model's size. ⚠️ Disclaimer: The "Vibes Test" These quantizations have NOT been formally tested for perplexity. They were compiled as an experiment to see how the model handles shifting bottlenecks. There is no guarantee that they are mathematically optimal or perform flawlessly. If they pass the vi…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy