🎨 Open PerfectBlend Open PerfectBlend is an open source reproduction of the instruction dataset introduced in the paper "The Perfect Blend: Redefining RLHF with Mixture of Judges". It's a solid general purpose instruction dataset with chat, math, code, and instruction following data. Data source Here is the list of the datasets used in this mix: Dataset Samples meta math/MetaMathQA 395,000 openbmb/UltraInteract sft 288,579 HuggingFaceH4/ultrachat 200k 207,865 microsoft/orca math word problems 200k 200,035 HuggingFaceH4/ultrafeedback binarized 187,405 theblackcat102/evol codealpaca v1 111,272 Post training Data Flywheel/AutoIF instruct 61k 61,492 mlabonne/lmsys arena human preference 55k sharegpt 57,362 The deduplication process removed 88.1k samples across all datasets. All of these datasets use either an Apache 2.0 or MIT license. Thanks to OpenBMB, MetaMath, Hugging Face, Microsoft, theblackcat102, Post training Data Flywheel, and LMSYS for the data! Comparison Here is the extract from the paper with the dataset mixture: There are two main differences with the dataset described in the paper: Instruction following data comes from another source because Meta didn't release their d…
Runs entirely in your browser via DuckDB-Wasm — this dataset's real data file is loaded once, then queried locally. Nothing is sent to a server.
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy