A multi-domain training dataset for query-conditioned extractive evidence
selection. Given a question and a passage, the task is to highlight the
verbatim substrings of the passage that support the answer.
Combines three sources covering distinct domains and annotation conventions:
source
domain
convention
annotator
rows (train / val)
ACL silver (this project)
NLP research papers
paragraph-scale
Qwen 3.6 35B (paragraph prompt)
20,916 / 2,319
RAGBench… See the full description on the dataset page:
https://huggingface.co/datasets/KRLabsOrg/verbatim-spans.