[!NOTE] For DeepSeek R1 0528 Qwen3 8B GGUFs, see here. Learn how to run DeepSeek R1 0528 correctly Read our Guide . See our collection for all versions of R1 including GGUF, 4 bit & 16 bit formats. Unsloth Dynamic 2.0 achieves superior accuracy & outperforms other leading quants. 🐋 DeepSeek R1 0528 Usage Guidelines Set the temperature between 0.5–0.7 (0.6 recommended) to reduce repetition and incoherence. Set Top P value of 0.95 (recommended) R1 0528 uses the same chat template as the original R1 model: For llama.cpp / GGUF inference, you should skip the BOS since it’ll auto add it: For complete detailed instructions, see our guide: unsloth.ai/blog/deepseek r1 0528 DeepSeek R1 0528 Model Card Paper Link 👁️ 1. Introduction The DeepSeek R1 model has undergone a minor version upgrade, with the current version being DeepSeek R1 0528. In the latest update, DeepSeek R1 has significantly improved its depth of reasoning and inference capabilities by leveraging increased computational resources and introducing algorithmic optimization mechanisms during post training. The model has demonstrated outstanding performance across various benchmark evaluations, including mathematics, programming…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy