Update [01/31/2024] We update the OpenAI Moderation API results for ToxicChat (0124) based on their updated moderation model on on Jan 25, 2024. [01/28/2024] We release an official T5 Large model trained on ToxicChat (toxicchat0124). Go and check it for you baseline comparision! [01/19/2024] We have a new version of ToxicChat (toxicchat0124)! Content This dataset contains toxicity annotations on 10K user prompts collected from the Vicuna online demo. We utilize a human AI collaborative annotation framework to guarantee the quality of annotation while maintaining a feasible annotation workload. The details of data collection, pre processing, and annotation can be found in our paper. We believe that ToxicChat can be a valuable resource to drive further advancements toward building a safe and healthy environment for user AI interactions. Version The version name is the update time of the dataset, e.g, 0124 means it is updated on Jan, 2024. We recommend using the latest version for training and evaluating a model. Please make sure the version of the data is the same when comparing different models. You can use the following code to specify the dataset version: toxicchat0124 Based on ve…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy