Fine tuned XLSR 53 large model for speech recognition in German Fine tuned facebook/wav2vec2 large xlsr 53 on German using the train and validation splits of Common Voice 6.1. When using this model, make sure that your speech input is sampled at 16kHz. This model has been fine tuned thanks to the GPU credits generously given by the OVHcloud :) The script used for training can be found here: https://github.com/jonatasgrosman/wav2vec2 sprint Usage The model can be used directly (without a language model) as follows... Using the HuggingSound library: Writing your own inference script: Reference Prediction ZIEHT EUCH BITTE DRAUSSEN DIE SCHUHE AUS. ZIEHT EUCH BITTE DRAUSSEN DIE SCHUHE AUS ES KOMMT ZUM SHOWDOWN IN GSTAAD. ES KOMMT ZUG STUNDEDAUTENESTERKT IHRE FOTOSTRECKEN ERSCHIENEN IN MODEMAGAZINEN WIE DER VOGUE, HARPER’S BAZAAR UND MARIE CLAIRE. IHRE FOTELSTRECKEN ERSCHIENEN MIT MODEMAGAZINEN WIE DER VALG AT DAS BASIN MA RIQUAIR FELIPE HAT EINE AUCH FÜR MONARCHEN UNGEWÖHNLICH LANGE TITELLISTE. FELIPPE HAT EINE AUCH FÜR MONACHEN UNGEWÖHNLICH LANGE TITELLISTE ER WURDE ZU EHREN DES REICHSKANZLERS OTTO VON BISMARCK ERRICHTET. ER WURDE ZU EHREN DES REICHSKANZLERS OTTO VON BISMARCK ERRICHTET…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy