Dataset Card for ToxiGen Table of Contents Dataset Description Dataset Summary Languages Dataset Structure Data Fields Additional Information Citation Information Sign up for Data Access To access ToxiGen, first fill out this form. Dataset Description Repository: https://github.com/microsoft/toxigen Paper: https://arxiv.org/abs/2203.09509 Point of Contact 1: Tom Hartvigsen Point of Contact 2: Saadia Gabriel Dataset Summary This dataset is for implicit hate speech detection. All instances were generated using GPT 3 and the methods described in our paper. Languages All text is written in English. Dataset Structure Data Fields We release TOXIGEN as a dataframe with the following fields: prompt is the prompt used for generation . generation is the TOXIGEN generated text. generation method denotes whether or not ALICE was used to generate the corresponding generation. If this value is ALICE, then ALICE was used, if it is TopK, then ALICE was not used. prompt label is the binary value indicating whether or not the prompt is toxic (1 is toxic, 0 is benign). group indicates the target group of the prompt. roberta prediction is the probability predicted by our corresponding RoBERTa model fo…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy