Technical Documentation for the Text-to-Video Dataset “VidData”
1. Introduction
This dataset contains 1006 annotated videos of everyday scenes, used for training and evaluating AI models in video generation and recognition. It is structured to meet the needs of Text-to-Video models and motion analysis.