Dataset Card for arxiv abstracts 2021 Table of Contents Dataset Description Dataset Summary Supported Tasks Languages Dataset Structure Data Instances Data Fields Data Splits Dataset Creation Curation Rationale Source Data Annotations Personal and Sensitive Information Considerations for Using the Data Social Impact of Dataset Discussion of Biases Other Known Limitations Additional Information Dataset Curators Licensing Information Citation Information Dataset Description Homepage: [Needs More Information] Repository: [Needs More Information] Paper: Clement et al., 2019, On the Use of ArXiv as a Dataset, https://arxiv.org/abs/1905.00075 Leaderboard: [Needs More Information] Point of Contact: Giancarlo Fissore Dataset Summary A dataset of metadata including title and abstract for all arXiv articles up to the end of 2021 (~2 million papers). Possible applications include trend analysis, paper recommender engines, category prediction, knowledge graph construction and semantic search interfaces. In contrast to arxiv dataset, this dataset doesn't include papers submitted to arXiv after 2021 and it doesn't require any external download. Supported Tasks and Leaderboards [Needs More Inform…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy