arXiv · 2406.04866
ComplexTempQA:A 100m Dataset for Complex Temporal Question Answering
Abstract
We introduce \textsc{ComplexTempQA},\footnote{Dataset and code available at: https://github.com/DataScienceUIBK/ComplexTempQA} a large-scale dataset consisting of over 100 million question-answer pairs designed to tackle the challenges in temporal question answering. \textsc{ComplexTempQA} significantly surpasses existing benchmarks in scale and scope. Utilizing Wikipedia and Wikidata, the dataset covers questions spanning over two decades and offers an unmatched scale. We introduce a new taxonomy that categorizes questions as \textit{attributes}, \textit{comparisons}, and \textit{counting} questions, revolving around events, entities, and time periods, respectively. A standout feature of \textsc{ComplexTempQA} is the high complexity of its questions, which demand reasoning capabilities for answering such as across-time comparison, temporal aggregation, and multi-hop reasoning involving temporal event ordering and entity recognition. Additionally, each question is accompanied by detailed metadata, including specific time scopes, allowing for comprehensive evaluation of temporal reasoning abilities of large language models.
Explore related subjects
Keep this discovery
Raphael Gruber, Abdelrahman Abdallah, Michael Färber, Adam Jatowt. 2024-06-07. ComplexTempQA:A 100m Dataset for Complex Temporal Question Answering. https://arxiv.org/abs/2406.04866
Cite the original work for its findings. Save a collection to share your selection of sources.