You are provided with a large number of Wikipedia comments which have been labeled by human raters for toxic behavior. The types of toxicity are:
You must create a model which predicts a probability of each type of toxicity for each comment.
train.csv - the training set, contains comments with their binary labels
test.csv - the test set, you must predict the toxicity… See the full description on the dataset page:
https://huggingface.co/datasets/thesofakillers/jigsaw-toxic-comment-classification-challenge.