Learn how to run DeepSeek R1 0528 correctly Read our Guide . See our collection for all versions of R1 including GGUF, 4 bit & 16 bit formats. Unsloth Dynamic 2.0 achieves superior accuracy & outperforms other leading quants. 🐋 DeepSeek R1 0528 Qwen3 8B Usage Guidelines Setting Non Thinking Mode Thinking Mode Temperature 0.7 0.6 Min P 0.0 0.0 Top P 0.8 0.95 TopK 20 20 Chat template/prompt format: For NON thinking mode, we purposely enclose and with nothing: For Thinking mode, DO NOT use greedy decoding, as it can lead to performance degradation and endless repetitions. For complete detailed instructions, see our guide: unsloth.ai/blog/deepseek r1 0528 DeepSeek R1 0528 Paper Link 👁️ 1. Introduction The DeepSeek R1 model has undergone a minor version upgrade, with the current version being DeepSeek R1 0528. In the latest update, DeepSeek R1 has significantly improved its depth of reasoning and inference capabilities by leveraging increased computational resources and introducing algorithmic optimization mechanisms during post training. The model has demonstrated outstanding performance across various benchmark evaluations, including mathematics, programming, and general logic. Its…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy