Dataset Card for WikiANN Table of Contents Dataset Description Dataset Summary Supported Tasks and Leaderboards Languages Dataset Structure Data Instances Data Fields Data Splits Dataset Creation Curation Rationale Source Data Annotations Personal and Sensitive Information Considerations for Using the Data Social Impact of Dataset Discussion of Biases Other Known Limitations Additional Information Dataset Curators Licensing Information Citation Information Contributions Dataset Description Homepage: Massively Multilingual Transfer for NER Repository: Massively Multilingual Transfer for NER Paper: The original datasets come from the Cross lingual name tagging and linking for 282 languages paper by Xiaoman Pan et al. (2018). This version corresponds to the balanced train, dev, and test splits of the original data from the Massively Multilingual Transfer for NER paper by Afshin Rahimi et al. (2019). Leaderboard: Point of Contact: Afshin Rahimi or Lewis Tunstall or Albert Villanova del Moral Dataset Summary WikiANN (sometimes called PAN X) is a multilingual named entity recognition dataset consisting of Wikipedia articles annotated with LOC (location), PER (person), and ORG (organisati…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy