Dutch TTS Dataset Complete Labeled A comprehensive Dutch text to speech dataset with 596,508 audio samples totaling 234GB of audio data. Quick Preview The default config shows a 100 row sample for the dataset viewer. To access the full dataset, use the full config. Dataset Description This dataset contains Dutch speech recordings with rich metadata including: Emotion labels (neutral, happy, sad, angry) Speaker IDs (239,388 unique speakers) Prosodic features (pitch mean/std, speaking rate) Audio quality metrics (SNR) Normalized text transcriptions Dataset Statistics Statistic Value Total samples 596,508 Total audio size ~234 GB Audio format WAV, 16kHz mono Unique speakers 239,388 Avg duration ~10 seconds Emotion Distribution neutral: 545,079 (91.4%) happy: 51,124 (8.6%) sad: 291 (0.05%) angry: 40 (0.01%) Loading the Dataset Features Feature Type Description audio Audio Audio waveform (16kHz) sampling rate int Always 16000 Hz duration float Duration in seconds text string Original transcription text normalized string Normalized transcription speaker id string Unique speaker identifier emotion string Emotion label emotion confidence float Confidence score (0 1) valence float Emotional…
Runs entirely in your browser via DuckDB-Wasm — this dataset's real data file is loaded once, then queried locally. Nothing is sent to a server.
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy