This dataset contains OCR results from images in Lukaszl/eu-law-ocr-dataset-pl-100 using DoTS.ocr, a compact 1.7B multilingual model.
Source Dataset: Lukaszl/eu-law-ocr-dataset-pl-100
Model: rednote-hilab/dots.ocr
Number of Samples: 100
Processing Time: 9.6 min
Processing Date: 2026-03-30 21:39 UTC
Image Column: image
Output Column: markdown
Dataset Split: train
Batch Size: 16
Prompt Mode: ocr… See the full description on the dataset page:
https://huggingface.co/datasets/Lukaszl/eu-law-ocr-dataset-pl-100-v1.