⚠️ Content warning. This dataset contains adversarial prompts that target nudity and violence safety mechanisms in text-to-image models. Many prompts reference real public figures. It is released for safety research only.
128 nudity + 47 violence natural-language prompts, each verified to jailbreak the four safe T2I models studied in the ICER paper: ESD, SLD-MAX, Receler, and AdvUnlearn (Stable Diffusion v1-4 backbone).
Generated by ICER, a… See the full description on the dataset page:
https://huggingface.co/datasets/zhiyichin/icer.