A Bilingual Multimodal Benchmark for Cognitively Informed Evaluation of Vision-Language Models
BloomBench is part of the Almieyar benchmarking series — the first cognitively human-grounded, bilingual (English–Arabic) multimodal benchmark for Vision-Language Models (VLMs). Grounded in Bloom's Taxonomy, it systematically evaluates six levels of cognition through carefully designed image–question–answer… See the full description on the dataset page:
https://huggingface.co/datasets/QCRI/BloomBench.