Temp-R1: A Unified Autonomous Agent for Complex Temporal KGQA via Reverse Curriculum Reinforcement Learning
Zhaoyan Gong, Zhiqiang Liu, Songze Li, Xiaoke Guo, Yuanxiang Liu, Xinle Deng, Zhizhen Liu, Lei Liang, Huajun Chen, Wen Zhang
Abstract
Temporal Knowledge Graph Question Answering (TKGQA) is inherently challenging, as it requires sophisticated reasoning over dynamic facts with multi-hop dependencies and complex temporal constraints. Existing methods rely on fixed workflows and expensive closed-source APIs, limiting flexibility and scalability. We propose Temp-R1, the first autonomous endto-end agent for TKGQA trained through reinforcement learning. To address cognitive overload in single-action reasoning, we expand the action space with specialized internal actions alongside external action. To prevent shortcut learning on simple questions, we introduce reverse curriculum learning that trains on difficult questions first, forcing the development of sophisticated reasoning before transferring to easier cases. Our 8B-parameter Temp-R1 achieves state-of-the-art performance on MUL-TITQ and TIMELINEKGQA, improving 19.8% over strong baselines on complex questions. Our work establishes a new paradigm for autonomous temporal reasoning agents. The code is available at https://github.com/zjukg/ Temp-R1 .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 8535f39e-1bc6-4e13-b6a7-a173742a7d04Cited by top-tier papers3
- ASTRA: Adaptive Semantic Tree Reasoning Architecture for Complex Table Question AnsweringXiaoke Guo, Songze Li, Zhiqiang Liu, Zhaoyan Gong et al.ACL 2026 · 3 citations
- CoG: Controllable Graph Reasoning via Relational Blueprints and Failure-Aware Refinement over Knowledge GraphsYuanxiang Liu, Songze Li, Xiaoke Guo, Zhaoyan Gong et al.ACL 2026 · 1 citation
- Collaboration of Fusion and Independence: Hypercomplex-driven Robust Multi-Modal Knowledge Graph CompletionZhiqiang Liu, Yichi Zhang, Mengshu Sun, Lei Liang et al.ACL 2026 · 1 citation
Builds on14
- Direct Preference Optimization: Your Language Model is Secretly a Reward ModelRafael Rafailov, Archit Sharma, Eric Mitchell, Christopher D. Manning et al.NeurIPS 2023 · 10,924 citations
- Improving Multi-hop Question Answering over Knowledge Graphs using Knowledge Base EmbeddingsApoorv Saxena, Aditay Tripathi, Partha P. TalukdarACL 2020 · 488 citations
- ReSearch: Learning to Reason with Search for LLMs via Reinforcement LearningMingyang Chen, Linzhuang Sun, Tianpeng Li, Haoze Sun et al.NeurIPS 2025 · 125 citations
- TempoQR: Temporal Question Reasoning over Knowledge GraphsCostas Mavromatis, Prasanna Lakkur Subramanyam, Vassilis N. Ioannidis, Adesoji Adeshina et al.AAAI 2022 · 77 citations
- Multi-granularity Temporal Question Answering over Knowledge GraphsZiyang Chen, Jinzhi Liao, Xiang ZhaoACL 2023 · 36 citations
Related papers
- KBQA-R1: Reinforcing Large Language Models for Knowledge Base Question AnsweringXin Sun, Zhongqi Chen, Xing Zheng, Bowen Song et al.ICML 2026 · 2 citations
- RTQA : Recursive Thinking for Complex Temporal Knowledge Graph Question Answering with Large Language ModelsZhaoyan Gong, Juan Li, Zhiqiang Liu, Lei Liang et al.EMNLP 2025 · 1 citation
- Reinforcement Learning Enhanced Muti-hop Reasoning for Temporal Knowledge Question AnsweringWuzhenghong Wen, Chao Xue, Su Pan, Yuwei Sun et al.AAAI 2026
- Temporal Evidence Chain for Temporal Knowledge Graph Question Answering with Large Language ModelsShihao Liu, Xiaofei Zhou, Bo Wang, Geyuan ZhangACL 2026
- TimeTraveler: Reinforcement Learning for Temporal Knowledge Graph ForecastingHaohai Sun, Jialun Zhong, Yunpu Ma, Zhen Han et al.EMNLP 2021 · 164 citations
