Accelerating the development of large-scale multi-modality models (LMMs) with lmms-eval
This is a formatted version of DocVQA. It is used in our lmms-eval pipeline to allow for one-click evaluations of large multi-modality models.
@article{mathew2020docvqa,
title={DocVQA: A Dataset for VQA on Document Images. CoRR abs/2007.00398 (2020)}… See the full description on the dataset page:
https://huggingface.co/datasets/lmms-lab-encoder/DocVQA.