Dataset Card for Orca DPO Pair Dataset Description This is a pre processed version of the OpenOrca dataset. The original OpenOrca dataset is a collection of augmented FLAN data that aligns, as best as possible, with the distributions outlined in the Orca paper. It has been instrumental in generating high performing preference tuned model checkpoints and serves as a valuable resource for all NLP researchers and developers! Dataset Summary The OrcaDPO Pair dataset is a subset of the OpenOrca dataset suitable for DPO preference tuning. The dataset is stored in parquet format with each entry using the following schema: : Data Splits The dataset consists of two splits, "train prefs" and "test prefs" : train prefs test prefs : : : : 12359 500 Usage To load the dataset, run: Languages The language of the data is primarily English. Dataset Creation Curation Rationale The dataset was created to provide a source of augmented text data for researchers and developers. The datapoints are intended primarily to provide an enhancement of the core FLAN Collection data which relies upon the detailed step by step reasoning capabilities of GPT 3.5 and GPT 4. This "reasoning trace" augmentation has dem…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy