Qwen3.5 27B Claude 4.6 Opus Reasoning Distilled v2 AWQ Base model: Jackrong/Qwen3.5 27B Claude 4.6 Opus Reasoning Distilled v2 This repo quantizes the model using data free quantization (no calibration dataset required). 【Dependencies / Installation】 As of 2026 03 30 , make sure your system has cuda12.8 installed. Then, create a fresh Python environment (e.g. python3.12 venv) and run: vLLM Official Guide 【vLLM Startup Command】 【Logs】 【Model Files】 File Size Last Updated 21GiB 2026 03 30 【Model Download】 【Overview】 🌟 Qwen3.5 27B Claude 4.6 Opus Reasoning Distilled v2 📢 Announcement v2 Update: Accuracy preserved: Matches base model on HumanEval ( 96.91% pass@1 ) Shorter reasoning: ~ 24% reduction in chain of thought length Higher efficiency: +31.6% more correct solutions per token ⚠️ Trade off: −1.24% on HumanEval+ −7.2% on MMLU Pro (Indicating reduced general knowledge reasoning performance) ⚠️Note: Due to the scope of SFT data and training focus, the model may underperform the base model on certain tasks requiring long context understanding or more complex multi step reasoning. The efficiency and accuracy results reported here are based solely on the HumanEval and HumanEval+ benc…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy