Qwopus GLM 18B Merged 🧪 A 64 layer frankenmerge of two of Jackrong's incredible Qwen3.5 9B finetunes, stacking all 32 layers from each to create an ~18B parameter model, then healed with a 1000 step LoRA fine tune to smooth the layer boundary. This was a fun experiment! A lot of people have been asking for something between Jackrong's 27B and 9B models — something that runs well on 12–16 GB GPUs. This frankenmerge is an attempt at filling that gap, and the results are surprisingly good. [!NOTE] Thanks to the creator of this model, @KyleHessling1 🙌 This is still an experimental model, so it may have quirks or issues. If you run into anything weird, or if you make something cool with it, reach out on X. Heal Fine Tune — It Works 🛠️ The raw frankenmerge had a known issue: garbled code output . Because two separately trained models were stacked at layer 32, structured output (code blocks, HTML, bracket matching) would occasionally come out malformed or hallucinated. We ran a 1000 step QLoRA heal fine tune using Jackrong's own training data to let gradients flow across the layer boundary — and the results are significant: HTML generation is now clean and production quality. We tested…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy