The only AI benchmark dataset covering LLM · VLM · Agent · Image · Video · Music in a single unified file.
ALL Bench Leaderboard aggregates and cross-verifies benchmark scores for 90+ AI models across 6 modalities. Every numerical score is tagged with a confidence level (cross-verified, single-source, or self-reported) and its original source. The dataset is designed for researchers, developers, and… See the full description on the dataset page:
https://huggingface.co/datasets/youssef3146/ALL-Bench-Leaderboard.