ASMR Archive Processed (WIP) Update (2026 04 03): This dataset has reached the Hugging Face Public Storage Limit. After contacting support, we were informed that the only option is to pay for a storage expansion. Consequently, updates to this dataset are now suspended. Work in Progress — expect breaking changes while the pipeline and data layout stabilize. This dataset contains ASMR audio data sourced from DeliberatorArchiver/asmr archive data 01 and DeliberatorArchiver/asmr archive data 02, which has undergone the following preprocessing steps: Preprocessing Steps 1. Low Quality Data Filtering : Audio files are filtered to remove low quality samples. This process checks for: Undesirable codecs (e.g., 8 bit PCM, ADPCM) Short durations (less than 12 seconds) Low sample rates (below 22,050 Hz) For lossy codecs, an insufficient bitrate (adjusted for stereo and higher sample rates) 2. Format Uniformization and Conversion : All audio files are converted to a uniform format: 44.1 kHz sample rate, 24 bit depth, stereo FLAC . (Note: Original mono tracks are also converted to stereo in this step.) 3. Background Noise Removal / Vocal Separation : Background noise is removed, and vocals are e…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy