Whisper Small Cantonese Alvin This model is a fine tuned version of openai/whisper small on the Cantonese language. It achieves a 7.93 CER (without punctuations), 9.72 CER (with punctuations) on Common Voice 16.0 Training and evaluation data For training, CantoMap: Winterstein, Grégoire, Tang, Carmen and Lai, Regine (2020) "CantoMap: a Hong Kong Cantonese MapTask Corpus", in Proceedings of The 12th Language Resources and Evaluation Conference, Marseille: European Language Resources Association, p. 2899 2906. Cantonse ASR: Yu, Tiezheng, Frieske, Rita, Xu, Peng, Cahyawijaya, Samuel, Yiu, Cheuk Tung, Lovenia, Holy, Dai, Wenliang, Barezi, Elham, Chen, Qifeng, Ma, Xiaojuan, Shi, Bertram, Fung, Pascale (2022) "Automatic Speech Recognition Datasets in Cantonese: A Survey and New Dataset", 2022. Link: https://arxiv.org/pdf/2201.02419.pdf Name of Hours Common Voice 16.0 zh HK Train 138 Common Voice 16.0 yue Train 85 Common Voice 17.0 yue Train 178 Cantonese ASR 72 CantoMap 23 Pseudo Labelled YouTube Data 438 For evaluation, Common Voice 16.0 yue Test set is used. Results CER (lower is better): 0.0972 down from 0.1073, 0.1581 in the previous versions CER (punctuations removed): 0.0793 GPU In…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy