This dataset is derived from danameyer/printed-test-lines and has been enriched with inference results.
This dataset contains 604 samples across 1 split(s).
Duplicate line statistics are calculated from the dataset key columns filename, region_id, and line_id, plus project_name when project metadata is… See the full description on the dataset page:
https://huggingface.co/datasets/danameyer/test-inference-full-upload-6523414f.