日本語はこちら
This dataset is made of soa-full.
soa-full is an CC-0 image dataset from Smithsonian Open Access. However, the dataset does not contain the image caption.
Therefore, we caption the images by Florence 2.
Research Vision & Language
Develop text-to-image model or image-to-text model.… See the full description on the dataset page:
https://huggingface.co/datasets/aipicasso/soa-full-florence2.