The MERIT Dataset is a multimodal dataset (image + text + layout) designed for training and benchmarking Large Language Models (LLMs) on Visually Rich Document Understanding (VrDU) tasks. It is a fully labeled synthetic dataset generated using our opensource pipeline available on GitHub. You can explore more details about the dataset and pipeline reading our paper.
AI faces some dynamic and technical issues that push⦠See the full description on the dataset page:
https://huggingface.co/datasets/de-Rodrigo/merit.