Dataset Card for tweet eval Table of Contents Dataset Description Dataset Summary Supported Tasks and Leaderboards Languages Dataset Structure Data Instances Data Fields Data Splits Dataset Creation Curation Rationale Source Data Annotations Personal and Sensitive Information Considerations for Using the Data Social Impact of Dataset Discussion of Biases Other Known Limitations Additional Information Dataset Curators Licensing Information Citation Information Contributions Dataset Description Homepage: [Needs More Information] Repository: GitHub Paper: EMNLP Paper Leaderboard: GitHub Leaderboard Point of Contact: [Needs More Information] Dataset Summary TweetEval consists of seven heterogenous tasks in Twitter, all framed as multi class tweet classification. The tasks include irony, hate, offensive, stance, emoji, emotion, and sentiment. All tasks have been unified into the same benchmark, with each dataset presented in the same format and with fixed training, validation and test splits. Supported Tasks and Leaderboards text classification : The dataset can be trained using a SentenceClassification model from HuggingFace transformers. Languages The text in the dataset is in English…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy