Project Page | Paper | GitHub
MA-EgoQA (Multi-Agent Egocentric Video Question Answering) is a benchmark designed to evaluate models on their ability to understand multiple long-horizon egocentric video streams simultaneously collected from embodied agents.
Built on the EgoLife dataset, it features 266 hours of multi-agent video where 6 people lived together for 7 days. The benchmark includes 1.7k… See the full description on the dataset page:
https://huggingface.co/datasets/KangsanKim71/MA-EgoQA.