wav2vec2 large xls r 300m Urdu This model is a fine tuned version of facebook/wav2vec2 xls r 300m on the common voice dataset. It achieves the following results on the evaluation set: Loss: 0.9889 Wer: 0.5607 Cer: 0.2370 Evaluation Commands 1. To evaluate on mozilla foundation/common voice 8 0 with split test Inference With LM Training hyperparameters The following hyperparameters were used during training: learning rate: 0.0001 train batch size: 32 eval batch size: 8 seed: 42 gradient accumulation steps: 2 total train batch size: 64 optimizer: Adam with betas=(0.9,0.999) and epsilon=1e 08 lr scheduler type: linear lr scheduler warmup steps: 1000 num epochs: 200 Training results Training Loss Epoch Step Validation Loss Wer Cer : : : : : : : : : : : : 3.6398 30.77 400 3.3517 1.0 1.0 2.9225 61.54 800 2.5123 1.0 0.8310 1.2568 92.31 1200 0.9699 0.6273 0.2575 0.8974 123.08 1600 0.9715 0.5888 0.2457 0.7151 153.85 2000 0.9984 0.5588 0.2353 0.6416 184.62 2400 0.9889 0.5607 0.2370 Framework versions Transformers 4.17.0.dev0 Pytorch 1.10.2+cu102 Datasets 1.18.2.dev0 Tokenizers 0.11.0 Eval results on Common Voice 8 "test" (WER): Without LM With LM (run ./eval.py ) 52.03 39.89
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy