As released in HuggingFace,
JavisBench is a comprehensive and challenging benchmark for evaluating text-to-audio-video generation models.It covers multiple aspects of generation quality, semantic alignment, and temporal synchrony, enabling thorough assessment in both controlled and real-world scenarios.
pip install… See the full description on the dataset page:
https://huggingface.co/datasets/JavisVerse/JavisBench.