Gemma4 26B A4B QAT Uncensored HauhauCS Balanced MTP Join the Discord for updates, roadmaps, projects, or just to chat. Gemma4 26B A4B (QAT) uncensored by HauhauCS. 0/465 Refusals About No changes to datasets or capabilities — fully functional, 100% of what the original authors intended, just without the refusals. Built from the official QAT weights, so the 4 bit quant stays close to full precision quality. Balanced The Balanced variant (recommended — 99%+ of users will be happy here) uses optimized full uncensoring tuned especially for agentic coding, reasoning, creative writing and reliability critical tasks. It reasons before answering and stays dependable and on instruction. An Aggressive variant, for cases where Balanced still deflects too much, after current testing is not required. ~35% faster with MTP Ships with an MTP (multi token prediction) draft head for speculative decoding — roughly 35% faster generation with identical output (the model verifies every drafted token, so quality is unchanged — pure speed). This release is tuned to pair well with the included MTP head. llama.cpp: Note: the MTP speedup was currently tested by me through llama.cpp ( llama server / llama cli…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy