STIPLAR is a real-world scene text image dataset containing Korean, Arabic, and Japanese text image pairs collected from MLT-2019 and web sources, designed for fine-tuning STELLAR on low-resource languages.
Dataset download location will be updated.
For Stage 2 training, we utilize STIPLAR, our newly proposed scene text image pairs of low-resource language and… See the full description on the dataset page:
https://huggingface.co/datasets/annms-stellar/stiplar.