Dataset Card for DAPO Math 17k Processed This is a processed version of BytedTsinghua SIA/DAPO Math 17k where we have: Deduplicated the prompts Reformatted the prompts and ground truth answers to be compatible with TRL's GRPO trainer We have also derived pure English and Chinese subsets. The full dataset processing logic can be found in create dataset.py. If you find this dataset useful in your work, please cite the original source with:
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy