Dataset Card for LibriTTS R LibriTTS R [1] is a sound quality improved version of the LibriTTS corpus (http://www.openslr.org/60/) which is a multi speaker English corpus of approximately 585 hours of read English speech at 24kHz sampling rate, published in 2019. Overview This is the LibriTTS R dataset, adapted for the datasets library. Usage Splits There are 7 splits (dots replace dashes from the original dataset, to comply with hf naming requirements): dev.clean dev.other test.clean test.other train.clean.100 train.clean.360 train.other.500 Configurations There are 3 configurations, each which limits the splits the load dataset() function will download. The default configuration is "all". "dev": only the "dev.clean" split (good for testing the dataset quickly) "clean": contains only "clean" splits "other": contains only "other" splits "all": contains only "all" splits Example Loading the clean config with only the train.clean.360 split. Streaming is also supported. Columns Example Row Dataset Details Dataset Description License: CC BY 4.0 Dataset Sources [optional] Homepage: https://www.openslr.org/141/ Paper: https://arxiv.org/abs/2305.18802 Citation
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy