Learn how to run DeepSeek R1 0528 correctly Read our Guide . See our collection for all versions of R1 including GGUF, 4 bit & 16 bit formats. Unsloth Dynamic 2.0 achieves superior accuracy & outperforms other leading quants. 🐋 DeepSeek R1 0528 Qwen3 8B Usage Guidelines For Ollama do ollama run hf.co/unsloth/DeepSeek R1 0528 Qwen3 8B GGUF:Q4 K XL it'll auto get the correct chat template and all settings Set the temperature between 0.5–0.7 (0.6 recommended) to reduce repetition and incoherence. Set Top P value of 0.95 (recommended) R1 0528 uses the same chat template as the original R1 model: For llama.cpp / GGUF inference, you should skip the BOS since it’ll auto add it: For complete detailed instructions, see our guide: unsloth.ai/blog/deepseek r1 0528 DeepSeek R1 0528 Model Card Paper Link 👁️ 1. Introduction The DeepSeek R1 model has undergone a minor version upgrade, with the current version being DeepSeek R1 0528. In the latest update, DeepSeek R1 has significantly improved its depth of reasoning and inference capabilities by leveraging increased computational resources and introducing algorithmic optimization mechanisms during post training. The model has demonstrated outsta…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy