Please visit the GitHub repository for other Myanmar Language datasets.
This is a reformatted version of alexbeatson/burmese_ocr_data converted into the native Hugging Face datasets format for easier loading and integration with modern OCR training pipelines.
This dataset contains Burmese text images and their corresponding ground truth text extracted from real-life documents, suitable for training Optical… See the full description on the dataset page:
https://huggingface.co/datasets/chuuhtetnaing/burmese_ocr_data_hf.