The dataset, randomly selected from MS-COCO, contains a total of 10K image-text pairs.
It could be used in the paper of Text-to-Image Diffusion Models can be Easily Backdoored through Multimodal Data Poisoning.
The dataset has been processed to conform to the format required by the GitHub repository BadT2I code.
If you find it useful in your research, please consider citing our paper:
@inproceedings{zhai2023text,
title={Text-to-image diffusion… See the full description on the dataset page:
https://huggingface.co/datasets/zsf/coco2014train_10k.