A counterfactual VQA dataset constructed using the CLEVR blender assets to procedurally generate both negative and normal counter factual VQA images and questions for the Multimodal Benchmark paper.
Dataset Structure
This repository contains counterfactual visual question answering data with:
Original images and counterfactual variants (modifications to test reasoning)
Questions for each image variant
Answer matrices showing how each image… See the full description on the dataset page: https://huggingface.co/datasets/scholo/MMB_dataset.