BannerBench: Benchmarking Vision Language Models for Multi-Ad Selection with Human Preferences
Dataset Summary
The BannerBench is designed to evaluate the ability of VLMs to identify the banner that best matches human preferences from a set of candidates.
Dataset Structure
The structure of the raw dataset is as follows:
{
"train": Dataset({
"features": [
'LPimage', 'image1', 'image2', 'image3', 'image4', 'image5'… See the full description on the dataset page: https://huggingface.co/datasets/cyberagent/BannerBench.