Qwable v2 GGUF GGUF quantizations of lordx64/Qwable v2 for use with llama.cpp and LM Studio. The base model is a reasoning distilled variant of Qwen3.6 35B A3B fine tuned to imitate the chain of thought style of Claude Opus 4.7. It thinks in explicit ... blocks before producing the final answer. Quant files See the file list for all available quant levels. Common choices: File Quant Approx size Use case .IQ4 XS.gguf IQ4 XS ~18 GB Smallest quant with good quality — default pick for LM Studio .Q4 K M.gguf Q4 K M ~21 GB Balanced quality / size .Q5 K M.gguf Q5 K M ~25 GB Higher quality .Q8 0.gguf Q8 0 ~35 GB Near lossless Running in llama.cpp Running in LM Studio Search for lordx64/Qwable v2 GGUF inside LM Studio's model browser and pick the quant that fits your RAM/VRAM. The model should appear automatically once HF indexes this repo. License Apache 2.0, inherited from the base model. See lordx64/Qwable v2 for training details, evaluations, and intended use.
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy