Wav2Vec2 Large XLSR 53 Kazakh Fine tuned facebook/wav2vec2 large xlsr 53 for Kazakh ASR using the Kazakh Speech Corpus v1.1 When using this model, make sure that your speech input is sampled at 16kHz. Usage The model can be used directly (without a language model) as follows: Evaluation The model can be evaluated as follows on the test set of Kazakh Speech Corpus v1.1. To evaluate, download the archive, untar and pass the path to data to get test dataset as below: Test Result : 19.65% Training The Kazakh Speech Corpus v1.1 train dataset was used for training.
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy