Per-polygon scene-text-recognition crops derived from
llm-jp/jawildtext.
Each row of the source dataset carries a polygons column with quadrilateral
text regions and their transcriptions. For every polygon we perspective-warp
the source image onto the rectified bounding rectangle, yielding a tight,
horizontally-aligned crop suitable for training/evaluating Japanese scene-text
recognition models.
Samples: 108 403
Shards: 22… See the full description on the dataset page:
https://huggingface.co/datasets/nagohachi/jawildtext_cropped.