Wav2vec2 xls r 1b for Finnish ASR This acoustic model is a fine tuned version of facebook/wav2vec2 xls r 1b for Finnish ASR. The model has been fine tuned with 275.6 hours of Finnish transcribed speech data. Wav2Vec2 XLS R was introduced in this paper and first released at this page. This repository also includes Finnish KenLM language model used in the decoding phase with the acoustic model. Note : this model is exactly the same as the aapot/wav2vec2 xlsr 1b finnish lm v2 model so that model has just been copied/moved to this Finnish NLP Hugging Face organization. Model description Wav2Vec2 XLS R is Facebook AI's large scale multilingual pretrained model for speech. It is pretrained on 436k hours of unlabeled speech, including VoxPopuli, MLS, CommonVoice, BABEL, and VoxLingua107. It uses the wav2vec 2.0 objective, in 128 languages. You can read more about the pretrained model from this blog and this paper. This model is fine tuned version of the pretrained model (1 billion parameter variant) for Finnish ASR. Intended uses & limitations You can use this model for Finnish ASR (speech to text) task. How to use Check the run finnish asr models.ipynb notebook in this repository for an…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy