Wav2Vec2 Large XLSR 53 Marathi Fine tuned facebook/wav2vec2 large xlsr 53 on Marathi using the Open SLR64 dataset. When using this model, make sure that your speech input is sampled at 16kHz. This data contains only female voices but the model works well for male voices too. Trained on Google Colab Pro on Tesla P100 16GB GPU. WER (Word Error Rate) on the Test Set : 12.70 % Usage The model can be used directly without a language model as follows, given that your dataset has Marathi actual text and path in folder columns: Evaluation Evaluated on 10% of the Marathi data on Open SLR 64. Training Train Test ratio was 90:10. The training notebook Colab link here. Training Config and Summary weights and biases run summary here
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy