Norwegian Wav2Vec2 Model 300M VoxRex Nynorsk This model is finetuned on top of feature extractor VoxRex model from the National Library of Sweden. The finetuned model achieves the following results on the test set with a 5 gram KenLM. The numbers in parentheses are the results without the language model: WER: 0.1222 (0.1537) CER: 0.0419 (0.0468) Model description This is one of several Wav2Vec models our team created during the 🤗 hosted Robust Speech Event. This is the complete list of our models and their final scores: Model Final WER : : : : NbAiLab/nb wav2vec2 1b bokmaal 6.33 NbAiLab/nb wav2vec2 300m bokmaal 7.03 NbAiLab/nb wav2vec2 1b nynorsk 11.32 NbAiLab/nb wav2vec2 300m nynorsk (this model) 12.22 Dataset In parallel with the event, the team also converted the Norwegian Parliamentary Speech Corpus (NPSC) to the NbAiLab/NPSC in 🤗 Dataset format and used that as the main source for training. Code We have released all the code developed during the event so that the Norwegian NLP community can build upon it when developing even better Norwegian ASR models. The finetuning of these models is not very computationally demanding. After following the instructions here, you should be…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy