Dataset Card for "blimp" Table of Contents Dataset Description Dataset Summary Supported Tasks and Leaderboards Languages Dataset Structure Data Instances Data Fields Data Splits Dataset Creation Curation Rationale Source Data Annotations Personal and Sensitive Information Considerations for Using the Data Social Impact of Dataset Discussion of Biases Other Known Limitations Additional Information Dataset Curators Licensing Information Citation Information Contributions Dataset Description Homepage: Repository: https://github.com/alexwarstadt/blimp Paper: BLiMP: The Benchmark of Linguistic Minimal Pairs for English Paper: https://arxiv.org/abs/1912.00582 Point of Contact: More Information Needed Size of downloaded dataset files: 29.58 MB Size of the generated dataset: 11.45 MB Total amount of disk used: 41.03 MB Dataset Summary BLiMP is a challenge set for evaluating what language models (LMs) know about major grammatical phenomena in English. BLiMP consists of 67 sub datasets, each containing 1000 minimal pairs isolating specific contrasts in syntax, morphology, or semantics. The data is automatically generated according to expert crafted grammars. Supported Tasks and Leaderboards…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy