Description: This repository contains the dataset for the D3PO method in this paper Using Human Feedback to Fine-tune Diffusion Models without Any Reward Model. The d3po_dataset file pertains to the image distortion experiment of the anything-v5 model.
The text2img_dataset comprises the images generated from the pretrained, preferred image fine-tuned, reward weighted fine-tuned and… See the full description on the dataset page:
https://huggingface.co/datasets/yangkaiSIGS/d3po_datasets.