Dataset Card for PAWS: Paraphrase Adversaries from Word Scrambling Table of Contents Dataset Description Dataset Summary Supported Tasks and Leaderboards Languages Dataset Structure Data Instances Data Fields Data Splits Dataset Creation Curation Rationale Source Data Annotations Personal and Sensitive Information Considerations for Using the Data Social Impact of Dataset Discussion of Biases Other Known Limitations Additional Information Dataset Curators Licensing Information Citation Information Contributions Dataset Description Homepage: PAWS Repository: PAWS Paper: PAWS: Paraphrase Adversaries from Word Scrambling Point of Contact: Yuan Zhang Dataset Summary PAWS: Paraphrase Adversaries from Word Scrambling This dataset contains 108,463 human labeled and 656k noisily labeled pairs that feature the importance of modeling structure, context, and word order information for the problem of paraphrase identification. The dataset has two subsets, one based on Wikipedia and the other one based on the Quora Question Pairs (QQP) dataset. For further details, see the accompanying paper: PAWS: Paraphrase Adversaries from Word Scrambling (https://arxiv.org/abs/1904.01130) PAWS QQP is not av…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy