This dataset features a collection of high-quality images sourced from Pexels and captioned using the Gemini-1.5-Flash API. This dataset is designed to provide accurate, detailed descriptions of various visual content, suitable for text-to-image tasks, training AI models, and more.
Gemini promt:
"Describe this image, for a text-to-image train to be accurate, max 74 tokens. (the common theme between these images is '{theme}'), prefer the use of ',' dont use '.' and there is no need to have a… See the full description on the dataset page:
https://huggingface.co/datasets/Pixel-Dust/Pexels_Gemini_capitoned.