Libriheavy Libriheavy: a 50,000 hours ASR corpus with punctuation casing and context. Libriheavy is a labeled version of Librilight. This uploaded version replaces the default Libri Light audio files with the highest quality available versions from librivox. In most cases, this consists an upgrade of the source audio from a 64kbps mp3 to a 128kbps mp3. Audio files are then re encoded using the Opus 68kbps codec to retain quality and reduce size. Homepage: https://github.com/k2 fsa/libriheavy License: apache 2.0 Configs Each dataset config exposes a single split named train . small ( train ): 509 hours of speech. 417 speakers averaging 1.22 hours per speaker. medium ( train ): 5042 hours of speech. 1531 speakers averaging 3.29 hours per speaker. large ( train ): 50794 hours of speech. 6736 speakers averaging 7.54 hours per speaker. dev ( train ): 22.3 hours of speech. 141 speakers averaging 0.16 hours per speaker. test clean ( train ): 10.5 hours of speech. 70 speakers averaging 0.15 hours per speaker. test other ( train ): 11.5 hours of speech. 72 speakers averaging 0.16 hours per speaker. test clean large ( train ): 107.5 hours of speech. 72 speakers averaging 1.49 hours per speak…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy