New! All the splits from the SemEval competition, including the test set, are now available in this page.
This repository contains the original 80 datasets used for the paper Question Answering over Tabular Data with DataBench:
A Large-Scale Empirical Evaluation of LLMs which appeared in LREC-COLING 2024 and the associated SemEval 2025 Task 8 competition.
Large Language Models (LLMs) are showing emerging abilities, and one of the latest recognized ones is⦠See the full description on the dataset page:
https://huggingface.co/datasets/cardiffnlp/databench.