PSI-VQA is a video question-answering benchmark built on the PSI 2.0 dataset, covering egocentric dashcam footage of pedestrian crossing scenarios. It is the OOD Test Set 2 for AI City Challenge 2026, Track 3: Driver Situation Awareness.
The dataset spans four complementary tasks, all unified under the tao-vl-reason-v1.0 schema (NVIDIA TAR Benchmark format). Each item pairs a short video clip with a question; the model must return a… See the full description on the dataset page:
https://huggingface.co/datasets/ise-ice-lab/PSI_VQA.