ComplexTempQA is a large-scale dataset designed for complex temporal question answering (TQA). It consists of over 100 million question-answer pairs, making it one of the most extensive datasets available for TQA. The dataset is generated using data from Wikipedia and Wikidata and spans questions over a period of 36 years (1987-2023).
Note: We have a smaller version consisting of questions from the time period 1987 until 2007.
Dataset Description… See the full description on the dataset page: https://huggingface.co/datasets/DataScienceUIBK/ComplexTempQA.