๐ง Qwen3.6 27B Claude Opus DeepSeek Distilled โ YaRN 1M Context A distilled Qwen3.6 27B GGUF with YaRN 4ร context extension to 1,048,576 tokens, optimized for local agentic reasoning, tool use, and long chain task execution. โ ๏ธ Sampling Parameters Please ensure temperature = 0.6 and top p = 0.95 when using this model. This model was trained and validated at these specific parameters. Both too high and too low temperatures cause problems: ๐ฅ Too high ( 0.6) โ Output becomes divergent as token probabilities flatten. This causes malformed tool calls, function name hallucinations, unstable parameter generation, and uncontrollable agent behavior. ๐ง Too low ( This is an important recommendation based on extensive real world testing. Please verify these parameters in your inference framework. ๐ข Highlights Area Score vs Qwen3.6 27B q4 k m BenchLocal 6 pack ๐ 86.5 +8.3 GPQA Diamond 198 ๐ฌ 83.84% +10.14% BugFind 15 ๐ 80 +20 ToolCall 15 ๐ง 97 +4 InstructFollow 15 ๐ 94 +17 StructOutput 15 ๐ 88 +11 MMLU 500 (5 shot) ๐ 91.80% ~tied (+0.2%) DataExtract 15 ๐ 81 2 Context window ๐๏ธ 1,048,576 4ร (262K โ 1M) ๐ Output speed : ~60 tok/s on A100 40GB ยท ~100 tok/s on RTX PRO 6000 (q4 k m + mtp=โฆ
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy