Fine tuned XLSR 53 large model for speech recognition in French Fine tuned facebook/wav2vec2 large xlsr 53 on French using the train and validation splits of Common Voice 6.1. When using this model, make sure that your speech input is sampled at 16kHz. This model has been fine tuned thanks to the GPU credits generously given by the OVHcloud :) The script used for training can be found here: https://github.com/jonatasgrosman/wav2vec2 sprint Usage The model can be used directly (without a language model) as follows... Using the HuggingSound library: Writing your own inference script: Reference Prediction "CE DERNIER A ÉVOLUÉ TOUT AU LONG DE L'HISTOIRE ROMAINE." CE DERNIER ÉVOLUÉ TOUT AU LONG DE L'HISTOIRE ROMAINE CE SITE CONTIENT QUATRE TOMBEAUX DE LA DYNASTIE ACHÉMÉNIDE ET SEPT DES SASSANIDES. CE SITE CONTIENT QUATRE TOMBEAUX DE LA DYNASTIE ASHEMÉNID ET SEPT DES SASANDNIDES "J'AI DIT QUE LES ACTEURS DE BOIS AVAIENT, SELON MOI, BEAUCOUP D'AVANTAGES SUR LES AUTRES." JAI DIT QUE LES ACTEURS DE BOIS AVAIENT SELON MOI BEAUCOUP DAVANTAGES SUR LES AUTRES LES PAYS BAS ONT REMPORTÉ TOUTES LES ÉDITIONS. LE PAYS BAS ON REMPORTÉ TOUTES LES ÉDITIONS IL Y A MAINTENANT UNE GARE ROUTIÈRE. IL AMNARD…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy