Whisper small singlish 122k. This model is a openai/whisper small, fine tuned on a subset (122k samples) of the National Speech Corpus. The following results on the evaluation set (43,788k samples) are reported: Loss: 0.171377 WER: 9.69 Model Details Model Description Developed by: jensenlwt Model type: automatic speech recognition License: MIT Finetuned from model: openai/whisper small Uses The model is intended as exploration exercise to develop better ASR model for Singapore English (singlish). The recommended audio usage for testing should be: 1. Involves local Singapore slang, dialect, names, and terms etc. 2. Involves Singaporean accent. Direct Use To use the model in an application, you can make use of transformers : Out of Scope Use Long form audio Broken Singlish (typically from older generation) Poor quality audio (audio samples are recorded in a controlled environment) Conversation (as the model is not trained on conversation) Training Details Training Data We made use of the National Speech Corpus for training. In specific, we made use of Part 2 – which is a series of audio samples of prompted read speech recordings that involves local named entities, slang, and dialect…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy