A dataset of 40 personalized subjects: Person (10), Pets (5), Landmarks (5), Objects (15), and Fiction Characters (5).
The dataset is divided into train and test splits. The number of images per subject varies from 10-20 images.
@inproceedings{yochameleon,
author = {Thao Nguyen and Krishna Kumar Singh and Jing Shi and Trung Bui and Yong Jae Lee and Yuheng Li},
title = {Yo\textquotesingle Chameleon: Personalized Vision and Language Generation},
year =… See the full description on the dataset page:
https://huggingface.co/datasets/adjjsk/YoLLaVA.