A better version of this model is available: Oriserve/Whisper Hindi2Hinglish Apex Whisper Hindi2Hinglish Prime: GITHUB LINK: github link SPEECH TO TEXT ARENA: Speech To Text Arena Table of Contents: Key Features Training Data Finetuning Usage Performance Overview Qualitative Performance Overview Quantitative Performance Overview Miscellaneous Key Features: 1. Hinglish as a language : Added ability to transcribe audio into spoken Hinglish language reducing chances of grammatical errors 2. Whisper Architecture : Based on the whisper architecture making it easy to use with the transformers package 3. Better Noise handling : The model is resistant to noise and thus does not return transcriptions for audios with just noise 4. Hallucination Mitigation : Minimizes transcription hallucinations to enhance accuracy. 5. Performance Increase : ~39% average performance increase versus pretrained model across benchmarking datasets Training: Data: Duration : A total of ~550 Hrs of noisy Indian accented Hindi data was used to finetune the model. Collection : Due to a lack of ASR ready hinglish datasets available, a specially curated proprietary dataset was used. Labelling : This data was then labe…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy