Fine tuned XLSR 53 large model for speech recognition in Persian Fine tuned facebook/wav2vec2 large xlsr 53 on Persian using the train and validation splits of Common Voice 6.1. When using this model, make sure that your speech input is sampled at 16kHz. This model has been fine tuned thanks to the GPU credits generously given by the OVHcloud :) The script used for training can be found here: https://github.com/jonatasgrosman/wav2vec2 sprint Usage The model can be used directly (without a language model) as follows... Using the HuggingSound library: Writing your own inference script: Reference Prediction از مهمونداری کنار بکشم از مهمانداری کنار بکشم برو از مهرداد بپرس. برو از ماقدعاد به پرس خب ، تو چیكار می كنی؟ خوب تو چیکار می کنی مسقط پایتخت عمان در عربی به معنای محل سقوط است مسقط پایتخت عمان در عربی به بعنای محل سقوط است آه، نه اصلاُ! اهنه اصلا توانست توانست قصیده فن شعر میگوید ای دوستان قصیده فن شعر میگوید ایدوستون دو استایل متفاوت دارین دوبوست داریل و متفاوت بری دو روز قبل از کریسمس ؟ اون مفتود پش پشش ساعت های کاری چیست؟ این توری که موشیکل خب Evaluation The model can be evaluated as follows on the Persian test data of Common Voice. Test Result : In the table below I report the…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy