Skip to content
Advertisement

Dataset Card for ToxiGen Sign up for Data Access To access ToxiGen, first fill out this form. Dataset Summary This dataset is for implicit hate speech detection. All instances were generated using GPT-3 and the methods described in our paper. Languages All text is written in English. Dataset Structure Data Fields We release TOXIGEN as a dataframe with the following fields: prompt is the prompt used for generation. generation is… See the full description on the dataset page:

Source: Hugging Face Hub (toxigen/toxigen-data). Metadata imported from the dataset’s Hub tags.

Advertisement