Qwen3.5 397B A17B Opus 4.6 Reasoning Uncensored GGUF The world's first reasoning enhanced uncensored 397B model. Abliterated + LoRA fine tuned on 12,842 high quality reasoning samples distilled from Anthropic's Opus 4.6 outputs. This is Stage 2 of the Qwen3.5 397B pipeline: Stage 1 — Abliterated (refusals removed), no fine tuning Stage 2 (this repo) — Abliterated + LoRA reasoning fine tune. Better chain of thought, deeper analysis, more structured problem solving 397B total parameters, 17B active per token (Mixture of Experts). Trained for 3,046 steps across 8×H200 GPUs. Final loss: 0.363 , accuracy: 90.2% . What's Different From Stage 1 Stage 1 (Abliterated Only) Stage 2 (This Model) Abliteration ✅ Custom pipeline ✅ Same pipeline Fine tuning None LoRA r=64, 134.7M trainable params Training data None 12,842 reasoning samples (Opus 4.6 distillation) Reasoning quality Base Qwen3.5 Enhanced chain of thought + structured analysis Thinking mode Default Trained with tags for explicit reasoning Final loss N/A 0.363 Final accuracy N/A 90.2% Training Details LoRA Configuration: Rank: 64, Alpha: 128 Target modules: self attn.{q,k,v,o} proj , shared expert.{gate,up,down} proj Trainable parame…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy