Wav2Vec2 Large XLSR 53 ml Fine tuned facebook/wav2vec2 large xlsr 53 on ml (Malayalam) using the Indic TTS Malayalam Speech Corpus (via Kaggle), Openslr Malayalam Speech Corpus, SMC Malayalam Speech Corpus and IIIT H Indic Speech Databases. The notebooks used to train model are available here. When using this model, make sure that your speech input is sampled at 16kHz. Usage The model can be used directly (without a language model) as follows: Evaluation The model can be evaluated as follows on the test data of combined custom dataset. For more details on dataset preparation, check the notebooks mentioned at the end of this file. Test Result (WER) : 28.43 % Training A combined dataset was created using Indic TTS Malayalam Speech Corpus (via Kaggle), Openslr Malayalam Speech Corpus, SMC Malayalam Speech Corpus and IIIT H Indic Speech Databases. The datasets were downloaded and was converted to HF Dataset format using this notebook The notebook used for training and evaluation can be found here
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy