Dataset Card for DataFlow-MM-ContextVQA
Dataset Summary
DataFlow-MM-ContextVQA is a large-scale synthetic multimodal dataset consisting of over 200,000 visual question–answer (VQA) instances. Each example pairs an image with a natural language question and an associated context document that contains the information required to derive the correct answer.
The dataset is designed to emphasize context-aware multimodal reasoning, where models must jointly leverage visual… See the full description on the dataset page: https://huggingface.co/datasets/OpenDCAI/dataflow-mm-context_vqa.