Corvus-OCR-Caption-Mini-Mix is a high-quality, compact image-caption dataset designed for training and evaluating image-to-text models. It is a carefully curated subset of the larger BLIP3o/BLIP3o-Pretrain-Long-Caption, optimized for mixed OCR and long-form captioning tasks.
Long-form natural language captions
OCR-heavy samples with scientific, mathematical, and document-style… See the full description on the dataset page:
https://huggingface.co/datasets/prithivMLmods/Corvus-OCR-Caption-Mini-Mix.