SuperGemma4 26B Uncensored Fast GGUF v2 The fast, uncensored llama.cpp build of the strongest SuperGemma text line. This release is for people who want three things together: a model that feels less censored than stock chat releases a model that is more capable than the raw base on practical text workloads a compact local GGUF that still serves quickly on Apple Silicon Why this build Uncensored chat behavior without forcing every prompt into coding mode Tuned from the strongest fast line instead of the raw base Neutral chat template baked into the GGUF to reduce prompt routing bugs Verified on Apple Silicon with clean general chat and coding responses Headline numbers Base model: google/gemma 4 26B A4B it Format: GGUF Q4 K M General Korean prompt speed: 222.0 tok/s Generation speed: 89.4 tok/s Derived from the verified SuperGemma Fast MLX line Why this build is appealing Carries the stronger Fast weights instead of the plain stock base Keeps general chat natural instead of routing everything into coding mode Preserves the uncensored release identity while staying useful on normal prompts Gives you a practical llama.cpp deployment target without losing the personality of the tuned l…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy