HUI German 51 Speakers Synthetic (Cleaned) Multispeaker German TTS dataset created and published by @dida 80b. The voices are derived from one reference recording per speaker from the HUI Audio Corpus German. The released audio here is newly synthesized with Qwen3 TTS; it is not a repackaging of the original HUI audio corpus. Dataset 51 synthetic speaker voices derived from one HUI reference recording per speaker ~102h synthetic audio Audio synthesized with Qwen3 TTS from German transcript text Clips segmented to max ~15s, 24 kHz mono WAV Validation status This release was checked locally after publication metadata repair: 30,858 referenced WAV files 0 missing referenced WAV files 0 unreferenced WAV files 51 speakers 24 kHz mono WAV throughout If the Hugging Face Dataset Viewer is unavailable for this repository, use the files directly via train list.txt , val list.txt , speaker map.txt , and the audio/ directory. Splits Split Samples Train 29,339 Validation 1,519 Total 30,858 (Initial release list: 30,859 — one outlier sample excluded, see below) Release Notes 1. ?? Encoding Artifact → ʊɾ (8,162 samples fixed) The initial published list files contained ?? artifacts where the IPA p…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy