Dataset Card for H4 Stack Exchange Preferences Dataset Dataset Description Homepage: https://archive.org/details/stackexchange Repository: (private for now) https://github.com/huggingface/h4 Point of Contact: Nathan Lambert, nathan@huggingface.co Size of downloaded dataset: 22.13 GB Number of instructions: 10,741,532 Dataset Summary This dataset contains questions and answers from the Stack Overflow Data Dump for the purpose of preference model training . Importantly, the questions have been filtered to fit the following criteria for preference models (following closely from Askell et al. 2021): have =2 answers . This data could also be used for instruction fine tuning and language model training. The questions are grouped with answers that are assigned a score corresponding to the Anthropic paper: Some important notes when using this dataset for preference model pretraining (PMP), which can be ignored for other uses: the data will likely need to be filtered more due to matching scores. see section 4.1 of Askel et al 2021 for instructions on using each pair of samples twice via the following binarization (for better pre training initialization): To see all the stackexchanges used i…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy