mT5 multilingual XLSum This repository contains the mT5 checkpoint finetuned on the 45 languages of XL Sum dataset. For finetuning details and scripts, see the paper and the official repository. Using this model in transformers (tested on 4.11.0.dev0) Benchmarks Scores on the XL Sum test sets are as follows: Language ROUGE 1 / ROUGE 2 / ROUGE L Amharic 20.0485 / 7.4111 / 18.0753 Arabic 34.9107 / 14.7937 / 29.1623 Azerbaijani 21.4227 / 9.5214 / 19.3331 Bengali 29.5653 / 12.1095 / 25.1315 Burmese 15.9626 / 5.1477 / 14.1819 Chinese (Simplified) 39.4071 / 17.7913 / 33.406 Chinese (Traditional) 37.1866 / 17.1432 / 31.6184 English 37.601 / 15.1536 / 29.8817 French 35.3398 / 16.1739 / 28.2041 Gujarati 21.9619 / 7.7417 / 19.86 Hausa 39.4375 / 17.6786 / 31.6667 Hindi 38.5882 / 16.8802 / 32.0132 Igbo 31.6148 / 10.1605 / 24.5309 Indonesian 37.0049 / 17.0181 / 30.7561 Japanese 48.1544 / 23.8482 / 37.3636 Kirundi 31.9907 / 14.3685 / 25.8305 Korean 23.6745 / 11.4478 / 22.3619 Kyrgyz 18.3751 / 7.9608 / 16.5033 Marathi 22.0141 / 9.5439 / 19.9208 Nepali 26.6547 / 10.2479 / 24.2847 Oromo 18.7025 / 6.1694 / 16.1862 Pashto 38.4743 / 15.5475 / 31.9065 Persian 36.9425 / 16.1934 / 30.0701 Pidgin 37.9574…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy