Dataset Card for VibraVox 👀 While waiting for the TooBigContentError issue to be resolved by the HuggingFace team, you can explore the dataset viewer of vibravox test which has exactly the same architecture. DATASET SUMMARY The VibraVox dataset is a general purpose audio dataset of french speech captured with body conduction transducers. This dataset can be used for various audio machine learning tasks : Automatic Speech Recognition (ASR) (Speech to Text , Speech to Phoneme) Audio Bandwidth Extension (BWE) Speaker Verification (SPKV) / identification Voice cloning etc ... Dataset usage VibraVox contains 4 subsets, corresponding to different situations tailored for specific tasks. To load a specific subset simply use the following command ( can be any of the following : , , , ): The dataset is also compatible with the streaming mode: Citations, links and details Homepage: For more information about the project, visit our project page on https://vibravox.cnam.fr Github repository: jhauret/vibravox : Source code for ASR, BWE and SPKV tasks using the Vibravox dataset pip installable Python package: pypi/project/vibravox : ASR, BWE and SPKV tasks using the Vibravox dataset Published pa…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy