[!tip]
Please note that you can play videos in the Dataset Viewer above by clicking on them. Even better, try out our interactive demo!
A dataset for studying domain-specific action recognition and in-context video learning in VLMs.
[!tip]
TLDR: grab the benchmark MP4s from videos/ and the benchmark JSONLs (i.e., the actual Q&As) from benchmarks/.
benchmarks/ contains JSONL files… See the full description on the dataset page:
https://huggingface.co/datasets/raivn/VideoNet.