Adversarial Benchmark of Cruelty to Animals (ABCA)
A benchmark for auditing how AI assistants respond to real-world user requests that carry animal-welfare implications. Each prompt is a naturalistic query — many in the user's original language — where a good answer must balance being genuinely helpful with avoiding the facilitation or encouragement of animal cruelty.
Status: pre-release. Scenarios and tier anchors may change without notice until v1.0.
What's in… See the full description on the dataset page: https://huggingface.co/datasets/CompassioninMachineLearning/abca.