EO-Bench: Embodied Reasoning Benchmark for Vision-Language Models
Overview
EO-Bench is a comprehensive benchmark designed to evaluate the embodied reasoning capabilities of vision-language models (VLMs) in robotics scenarios. This benchmark is part of the EO-1 project, which develops unified embodied foundation models for general robot control.
The benchmark assesses model performance across 12 distinct embodied reasoning categories, covering trajectory… See the full description on the dataset page: https://huggingface.co/datasets/IPEC-COMMUNITY/EO-Bench.