Czech wav2vec2 xls r 300m cs 250 This model is a fine tuned version of facebook/wav2vec2 xls r 300m on the common voice 8.0 dataset as well as other datasets listed below. It achieves the following results on the evaluation set: Loss: 0.1271 Wer: 0.1475 Cer: 0.0329 The eval.py script results using a LM are: WER: 0.07274312090176113 CER: 0.021207369275558875 Model description Fine tuned facebook/wav2vec2 large xlsr 53 on Czech using the Common Voice dataset. When using this model, make sure that your speech input is sampled at 16kHz. The model can be used directly (without a language model) as follows: Evaluation The model can be evaluated using the attached eval.py script: Training and evaluation data The Common Voice 8.0 train and validation datasets were used for training, as well as the following datasets: Šmídl, Luboš and Pražák, Aleš, 2013, OVM – Otázky Václava Moravce, LINDAT/CLARIAH CZ digital library at the Institute of Formal and Applied Linguistics (ÚFAL), Faculty of Mathematics and Physics, Charles University, http://hdl.handle.net/11858/00 097C 0000 000D EC98 3. Pražák, Aleš and Šmídl, Luboš, 2012, Czech Parliament Meetings, LINDAT/CLARIAH CZ digital library at the Inst…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy