Dataset Card for "trec" Table of Contents Dataset Description Dataset Summary Supported Tasks and Leaderboards Languages Dataset Structure Data Instances Data Fields Data Splits Dataset Creation Curation Rationale Source Data Annotations Personal and Sensitive Information Considerations for Using the Data Social Impact of Dataset Discussion of Biases Other Known Limitations Additional Information Dataset Curators Licensing Information Citation Information Contributions Dataset Description Homepage: https://cogcomp.seas.upenn.edu/Data/QA/QC/ Repository: More Information Needed Paper: More Information Needed Point of Contact: More Information Needed Size of downloaded dataset files: 0.36 MB Size of the generated dataset: 0.41 MB Total amount of disk used: 0.78 MB Dataset Summary The Text REtrieval Conference (TREC) Question Classification dataset contains 5500 labeled questions in training set and another 500 for test set. The dataset has 6 coarse class labels and 50 fine class labels. Average length of each sentence is 10, vocabulary size of 8700. Data are collected from four sources: 4,500 English questions published by USC (Hovy et al., 2001), about 500 manually constructed questi…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy