This dataset was created synthetically by Tanaos with the Artifex Python library.
The dataset is designed to train and evaluate guardrail systems — models that detect, classify, or filter unsafe, harmful or potentially dangerous content. It can be used to train moderation models or integrate LLM safety filters for applications like chatbots, content generation, and user-facing AI systems.
Our flagship guardrail model, tanaos-guardrail-v2… See the full description on the dataset page:
https://huggingface.co/datasets/tanaos/synthetic-guardrail-dataset-v2.