StreamingBench: Assessing the Gap for MLLMs to Achieve Streaming Video Understanding
๐ Project Page |
๐ arXiv Paper |
๐ฆ Dataset |
๐ Leaderboard
StreamingBench evaluates Multimodal Large Language Models (MLLMs) in real-time, streaming video understanding tasks. ๐
[NEW! 2025.05.15] ๐ฅ: Seed1.5-VL achieved ALL model SOTA with a score of 82.80 on the Proactive Output.
[NEW! 2025.03.17] โญ: ViSpeeker achieved Open-Source SOTA with a score of 61.60 on theโฆ See the full description on the dataset page: https://huggingface.co/datasets/mjuicem/StreamingBench.