The evaluation dataset data/samples-1680.jsonl.gz is the test set used in this paper.
Each line contains information about one sample in a JSON object and each sample is labeled according to our taxonomy. The category label is a binary flag, but if it does not include in the JSON, it means we do not know the label.
sexual
S
Content meant to arouse sexual… See the full description on the dataset page:
https://huggingface.co/datasets/mmathys/openai-moderation-api-evaluation.