Portable, sharded text-to-image training stages used by the VERB project. Each
stage contains original image bytes, SenseNova-compatible JSONL annotations,
a portable manifest, a file index, and per-shard SHA-256 metadata.
This public repository supports reproducible research. The packaged sources
do not share one uniform VERB license; users must follow each upstream source's
terms and review redistribution requirements before republishing the data.… See the full description on the dataset page:
https://huggingface.co/datasets/Haoruili46/VERB_Data.