⚡ Each donation = another big MoE quantized I host 25+ free APEX MoE quantizations as independent research. My only local hardware is an NVIDIA DGX Spark (122 GB unified memory) — enough for ~30 50B class MoEs, but bigger ones (200B+) require rented compute on H100/H200/Blackwell, typically $20 100 per quant. If APEX quants are useful to you, your support directly funds those bigger runs. 🎉 Patreon (Monthly) ☕ Buy Me a Coffee ⭐ GitHub Sponsors 💚 Big thanks to Hugging Face for generously donating additional storage — much appreciated. Gemma 4 26B A4B Heretic (Abliterated) APEX GGUF APEX (Adaptive Precision for EXpert Models) quantizations of gemma 4 26B A4B it heretic — an abliterated (uncensored) version of Gemma 4, created with the Heretic tool (v1.2.0) using Arbitrary Rank Ablation (ARA) on layers 10 30 to reduce refusals while preserving capabilities (KL divergence 0.0499 from original). Brought to you by the LocalAI team APEX Project Technical Report Benchmark Results Benchmarks coming soon. For reference APEX benchmarks on the Qwen3.5 35B A3B architecture, see mudler/Qwen3.5 35B A3B APEX GGUF. Available Files File Profile Size Best For gemma 4 26B…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy