This dataset contains expert-annotated multimodal interaction events for VideoEduBench, a diagnostic benchmark for evaluating multimodal large language models (MLLMs) on long-horizon classroom interaction understanding.
📄 Paper: VideoEduBench: Diagnostic Benchmarking of Video Agents for Long-Horizon Classroom Interaction Understanding (KDD 2026 Undergraduate Consortium)
🔗 Code repository:
https://github.com/tinaxie123/VideoDR-Benchmark