Multilingual Speech Recognition for Indonesian Languages This is the model built for the project Multilingual Speech Recognition for Indonesian Languages. It is a fine tuned facebook/wav2vec2 large xlsr 53 model on the Indonesian Common Voice dataset, High quality TTS data for Javanese SLR41, and High quality TTS data for Sundanese SLR44 datasets. We also provide a live demo to test the model. When using this model, make sure that your speech input is sampled at 16kHz. Usage The model can be used directly (without a language model) as follows: Evaluation The model can be evaluated as follows on the Indonesian test data of Common Voice. Test Result : 11.57 % Training The Common Voice train , validation , and ... datasets were used for training as well as ... and ... TODO The script used for training can be found here (will be available soon)
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy