Dataset Card for "xtreme" Dataset Summary The Cross lingual Natural Language Inference (XNLI) corpus is a crowd sourced collection of 5,000 test and 2,500 dev pairs for the MultiNLI corpus. The pairs are annotated with textual entailment and translated into 14 languages: French, Spanish, German, Greek, Bulgarian, Russian, Turkish, Arabic, Vietnamese, Thai, Chinese, Hindi, Swahili and Urdu. This results in 112.5k annotated pairs. Each premise can be associated with… See the full description on the dataset page: https://huggingface.co/datasets/google/xtreme.
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy