Merged Video Benchmark Captions and Semantic Embeddings
This dataset contains merged semantic caption segments and aligned text embeddings for four video benchmarks from a local DynamicvideoAgentRL/LVU cache. It includes captions and embeddings only, not source videos.
Files
captions.parquet: one row per caption segment.
semantic_vectors.float32.npy: NumPy array with shape 244909 x 3072; row i matches captions.parquet row where row_id == i.… See the full description on the dataset page: https://huggingface.co/datasets/CewEhao/videlseal_eval_datasets.