Qwen3.6 27B AEON Ultimate Uncensored GGUF This repository contains GGUF quantizations of the heavily fine tuned and uncensored AEON 7/Qwen3.6 27B AEON Ultimate Uncensored model. These quantizations were generated using a custom compiled llama.cpp 🧠 Advanced Architecture Preservation Unlike standard quantization, these models were generated using a high precision pipeline: Output Tensor Preservation: We utilized leave output tensor during quantization. This ensures the model's final projection head remains in FP16 precision, preventing the "numerical noise" that typically degrades the reasoning capabilities of smaller quantizations. SSM Weight Fidelity: The 1D SSM routing weights (alpha/beta) were preserved to maintain the model's complex long range memory and selective state dynamics. Native Reasoning: The model is equipped with a native reasoning trigger. By utilizing the built in block, the model can perform multi step logical planning before providing a final response. 🛠️ Key Fixes & Optimizations Applied Tokenizer Hash Bypass: Patched the custom BPE hash check to ensure full compatibility with modern llama.cpp inference engines. Native ChatML Injection: The tokenizer config.j…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy