🚀 Qwen3.6 14B A3B VibeForged v2 GGUF Welcome to the highly optimized, quantized version of tvall43/Qwen3.6 14B A3B VibeForged v2 ! This repository contains various llama.cpp GGUF formats to ensure this beast runs smoothly on your hardware, whether you're maxing out a high end GPU or squeezing every last drop of inference out of a modest laptop. 🧠 About the Model This is the fully repaired and fine tuned version of a pruned Qwen3.6 35B A3B heretic . Originally suffering from a bit of "brain damage" after being pruned down to 14B parameters via REAP, it was brought back to life by the user's trusty AI agent, Steve . Through an intensive vibetuning pass on thousands of real world "vibecoding" sessions extracted from the user's OpenCode SQLite database (and combined with Evol Instruct code data), Steve orchestrated a QLoRA fine tune that enforces strict ... XML reasoning boundaries and clean JSON tool calling syntax, resulting in this incredibly punchy and capable 14B MoE reasoning model. This is a multimodal model! We've also brought over the original multimodal projector ( mmproj ) files from the base heretic model so you can continue using its vision capabilities. 📦 Available For…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy