VLM-as-judge pairwise evaluation of OCR models on Polish EU law and legal-style documents. Results depend strongly on document type, so this should be read as a document-specific benchmark rather than a universal OCR ranking.
This benchmark focuses on dense Polish legal text derived from EU law materials, including multi-paragraph pages with small fonts, numbered sections, formal structure, references, and list-heavy… See the full description on the dataset page:
https://huggingface.co/datasets/Lukaszl/eu-law-ocr-dataset-pl-100-v1-results.