This dataset is part of the training data for the CVPR Workshop Metafood 2025 (MTF 2025) Dishcovery VLM Challenge. It consists of image-text pairs where the images are synthetically generated using Stable Diffusion 2.1.
The dataset has been carefully curated using the Precision at Scale: Domain-Specific Datasets On-Demand method, ensuring high relevance and quality for domain-specific tasks.
Associated… See the full description on the dataset page: https://huggingface.co/datasets/jesusmolrdv/MTF25-VLM-Challenge-Dataset-Synth.