OpenR1 Math 220k Dataset description OpenR1 Math 220k is a large scale dataset for mathematical reasoning. It consists of 220k math problems with two to four reasoning traces generated by DeepSeek R1 for problems from NuminaMath 1.5. The traces were verified using Math Verify for most samples and Llama 3.3 70B Instruct as a judge for 12% of the samples, and each problem contains at least one reasoning trace with a correct answer. The dataset consists of two splits: default with 94k problems and that achieves the best performance after SFT. extended with 131k samples where we add data sources like cn k12 . This provides more reasoning traces, but we found that the performance after SFT to be lower than the default subset, likely because the questions from cn k12 are less difficult than other sources. You can load the dataset as follows: Dataset curation To build OpenR1 Math 220k, we prompt DeepSeek R1 model to generate solutions for 400k problems from NuminaMath 1.5 using SGLang, the generation code is available here. We follow the model card’s recommended generation parameters and prepend the following instruction to the user prompt: "Please reason step by step, and put your final…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy