This is the Repository for CC-OCR Benchmark.
Dataset and evaluation code for the Paper "CC-OCR: A Comprehensive and Challenging OCR Benchmark for Evaluating Large Multimodal Models in Literacy".
Here is hosting the tsv version of CC-OCR data, which is used for evaluation in VLMEvalKit. Please refer to our GitHub for more information.
Model… See the full description on the dataset page:
https://huggingface.co/datasets/wulipc/CC-OCR.