A mini subset of the DocVQA dataset with 500 randomly selected question-answer pairs for document visual question answering evaluation.
Total Samples: 500 QA pairs
Source: DocVQA validation set
Task: Document Visual Question Answering
Image Format: PNG (extracted from parquet-embedded images)
image: Document image
question: Question about the document
answers: List of valid… See the full description on the dataset page:
https://huggingface.co/datasets/kenza-ily/docvqa_disco.