Living in the Moment: Can Large Language Models Grasp Co-Temporal Reasoning?
Proceedings of the 62nd Annual Meeting of the Association for Computational Linguistics (ACL), 2024
Abstract
This work introduces CoTempQA, a co-temporal question answering benchmark with four scenarios for evaluating reasoning about concurrent and interconnected events. Experiments reveal a substantial gap between current language models and human performance on co-temporal reasoning.
