Drive-P2D: A Progressive Perception-to-Decision Benchmark for VLMs in Autonomous Driving
Drive-P2D is a progressive perception-to-decision benchmark for evaluating Vision-Language Models (VLMs) across object-level perception, scene understanding, and driving decision-making in autonomous driving. It contains 6,650 questions across three levels: Object, Scene, and… See the full description on the dataset page:
https://huggingface.co/datasets/ColamentosZJU/Drive-P2D.