Automatic Speech Recognition for Ethiopian Languages πͺπΉ arXiv π [ preprint ] βοΈ Model Description Ethio ASR is a suite of multilingual Automatic Speech Recognition (ASR) models that support five Ethiopian languages: Amharic, Tigrinya, Afaan Oromo, Sidama, and Wolaytta. The ASR model in this repo is based on the wav2vec2βbert 2.0 pre trained model by fine tuning it on the WAXAL Speech Dataset. Developed by: Ethio ASR Team Task: Speech Recognition (ASR) and Language Identification (LID) Languages: Amharic, Tigrinya, Afaan Oromo, Sidama, and Wolaytta License: CC BY 4.0 Finetuned from: facebook/w2v bert 2.0 π Evaluation on WAXAL Test Set π ASR model in this HF repo Model Params Amharic Tigrinya Oromo Wolaytta Sidaama Avg. Ethio ASR (afrihubert) 92M 30.95 42.42 27.57 40.44 34.02 35.08 Ethio ASR (mms 300) 300M 30.19 41.62 26.41 39.10 32.66 33.99 Ethio ASR (mms 1b) 1B 26.14 37.63 23.69 37.51 31.02 31.20 Ethio ASR (w2v bert 2.0) π 600M 22.92 35.22 24.44 38.19 31.65 30.48 π Examples Language Audio Human Transcription ASR Transcription 1 Oromo Suuraan asii gaditti argaa jirru kun lafa gurgurtaa kuduraa fi muduraa dha. Kuduraa fi muduraan nyaataaf kan baay'ee namatti toluudha. Nyaachuuβ¦
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy