ViInfographicVQA is a Vietnamese Visual Question Answering (VQA) benchmark for infographic understanding.It evaluates models’ ability to read, reason, and synthesize information from data-rich, layout-heavy visuals that mix text, charts, maps, and design elements.
Two settings are provided:
Single-image VQA – questions answered from one infographic.
Multi-image VQA – questions requiring reasoning across multiple, semantically related… See the full description on the dataset page: https://huggingface.co/datasets/duytranus/ViInfographicVQA.