BiomedSQL
GitHub
Paper
Dataset Summary
BiomedSQL is a text-to-SQL benchmark designed to evaluate Large Language Models (LLMs) on scientific tabular reasoning tasks. It consists of curated question-SQL query-answer triples covering a variety of biomedical and SQL reasoning types. The benchmark challenges models to apply implicit
scientific criteria rather than simply translating syntax.
Repository Organization
benchmark_data: contains the question-SQL query-answer triples.
Please note that you… See the full description on the dataset page:
https://huggingface.co/datasets/NIH-CARD/BiomedSQL.