Dataset used to train Three Kingdoms text to image model
The original images were obtained from
https://kongming.net/ and captioned with the pre-trained BLIP model.
For each row the dataset contains image and text keys. image is a varying size PIL jpeg, and text is the accompanying text caption. Only a train split is provided.
a man with a feather on… See the full description on the dataset page:
https://huggingface.co/datasets/wx44wx/three-kingdoms-blip-captions.