The Orca Moderation Corpus is a public dataset designed for text content moderation tasks. It contains examples of profanity, hate speech, violent language, and offensive content across multiple contexts. This dataset can be used to train, evaluate, and benchmark models for detecting toxic and harmful text.
⚠️ Warning: This dataset contains explicit, offensive, and potentially disturbing content. Use responsibly and in compliance with ethical AI guidelines.… See the full description on the dataset page:
https://huggingface.co/datasets/MrRamyg/orca-moderation-corpus.