This dataset is a toy format-checking dataset for the LongLive2.0 release
code. It is intended to help users verify AR diffusion training, DMD
distillation, and prompt formatting before preparing a larger dataset.
Dataset placeholder:
https://huggingface.co/datasets/Efficient-Large-Model/LongLive2-Toy-Dataset
ar_training/: paired video/caption data for AR… See the full description on the dataset page:
https://huggingface.co/datasets/Efficient-Large-Model/LongLive2.0-Toy-Dataset.