Our dataset consists of 906 diverse QA pairs in German and English.
The dataset is extractive, i.e., answers are given as sentence indices (breaking at the newline character \n).
Questions are automatically generated using an LLM.
The answers are manually annotated using voluntary crowdsourcing.
Repository: More Information Needed
Paper: