Fine tuned XLSR 53 large model for speech recognition in Portuguese Fine tuned facebook/wav2vec2 large xlsr 53 on Portuguese using the train and validation splits of Common Voice 6.1. When using this model, make sure that your speech input is sampled at 16kHz. This model has been fine tuned thanks to the GPU credits generously given by the OVHcloud :) The script used for training can be found here: https://github.com/jonatasgrosman/wav2vec2 sprint Usage The model can be used directly (without a language model) as follows... Using the HuggingSound library: Writing your own inference script: Reference Prediction NEM O RADAR NEM OS OUTROS INSTRUMENTOS DETECTARAM O BOMBARDEIRO STEALTH. NEMHUM VADAN OS OLTWES INSTRUMENTOS DE TTÉÃN UM BOMBERDEIRO OSTER PEDIR DINHEIRO EMPRESTADO ÀS PESSOAS DA ALDEIA E DIR ENGINHEIRO EMPRESTAR AS PESSOAS DA ALDEIA OITO OITO TRANCÁ LOS TRANCAUVOS REALIZAR UMA INVESTIGAÇÃO PARA RESOLVER O PROBLEMA REALIZAR UMA INVESTIGAÇÃO PARA RESOLVER O PROBLEMA O YOUTUBE AINDA É A MELHOR PLATAFORMA DE VÍDEOS. YOUTUBE AINDA É A MELHOR PLATAFOMA DE VÍDEOS MENINA E MENINO BEIJANDO NAS SOMBRAS MENINA E MENINO BEIJANDO NAS SOMBRAS EU SOU O SENHOR EU SOU O SENHOR DUAS MULHERES…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy