This public repository contains 5,000 text-to-video synthetic clips only. It
does not contain TalkVid/YouTube source videos or image-to-video clips
initialized from real-person reference images.
data_manifest.csv records media identity and probe metadata. Its
source_path values are… See the full description on the dataset page:
https://huggingface.co/datasets/hugging4chang/gfvc-vace-synthetic-t2v-data.