Dataset for the paper: WebUIBench: A Comprehensive Benchmark for Evaluating Multimodal Large Language Models in WebUI-to-Code
🏠 Homepage | 📖 arXiv
We introduce WebUIBench, a large-scale and comprehensive benchmark designed to evaluate the WebUI-to-Code capabilities of Multimodal Large Language Models (MLLMs). WebUIBench comprises over 21K question-answer pairs derived from more than 0.7K real-world websites, encompassing 9 distinct subtasks. We… See the full description on the dataset page:
https://huggingface.co/datasets/Tele-AI-MAIL/WebUIBench.