ECONET: Effective Continual Pretraining of Language Models for Event Temporal Reasoning
Rujun Han, Xiang Ren, Nanyun Peng
摘要
While pre-trained language models (PTLMs) have achieved noticeable success on many NLP tasks, they still struggle for tasks that require event temporal reasoning, which is essential for event-centric applications. We present a continual pre-training approach that equips PTLMs with targeted knowledge about event temporal relations. We design self-supervised learning objectives to recover masked-out event and temporal indicators and to discriminate sentences from their corrupted counterparts (where event or temporal indicators got replaced). By further pre-training a PTLM with these objectives jointly, we reinforce its attention to event and temporal information, yielding enhanced capability on event temporal reasoning. This Effective CONtinual pre-training framework for Event Temporal reasoning (ECONET) improves the PTLMs' fine-tuning performances across five relation extraction and question answering tasks and achieves new or on-par state-of-the-art performances in most of our downstream tasks. 1
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper14
- Large Language Models-guided Dynamic Adaptation for Temporal Knowledge Graph ReasoningJiapu Wang, Kai Sun, Linhao Luo, Wei Wei 等NeurIPS 2024 · 被引用 82 次
- Improving Time Sensitivity for Question Answering over Temporal Knowledge GraphsChao Shang, Guangtao Wang, Peng Qi, Jing HuangACL 2022 · 被引用 55 次
- LEMON: Lossless model expansionYite Wang, Jiahao Su, Hanlin Lu, Cong Xie 等ICLR 2024 · 被引用 25 次
- More than Classification: A Unified Framework for Event Temporal Relation ExtractionQuzhe Huang, Yutong Hu, Shengqi Zhu, Yansong Feng 等ACL 2023 · 被引用 14 次
- TimeBench: A Comprehensive Evaluation of Temporal Reasoning Abilities in Large Language ModelsZheng Chu, Jingchang Chen, Qianglong Chen, Weijiang Yu 等ACL 2024 · 被引用 12 次
它引用的顶会 Paper9
- ELECTRA: Pre-training Text Encoders as Discriminators Rather Than GeneratorsKevin Clark, Minh-Thang Luong, Quoc V. Le, Christopher D. ManningICLR 2020 · 被引用 541 次
- TANDA: Transfer and Adapt Pre-Trained Transformer Models for Answer Sentence SelectionSiddhant Garg, Thuy Vu, Alessandro MoschittiAAAI 2020 · 被引用 229 次
- Content Planning for Neural Story Generation with Aristotelian RescoringSeraphina Goldfarb-Tarrant, Tuhin Chakrabarty, Ralph M. Weischedel, Nanyun PengEMNLP 2020 · 被引用 106 次
- Joint Constrained Learning for Event-Event Relation ExtractionHaoyu Wang, Muhao Chen, Hongming Zhang, Dan RothEMNLP 2020 · 被引用 105 次
- TORQUE: A Reading Comprehension Dataset of Temporal Ordering QuestionsQiang Ning, Hao Wu, Rujun Han, Nanyun Peng 等EMNLP 2020 · 被引用 79 次
相关 Paper
- Pre-training Text-to-Text Transformers for Concept-centric Common SenseWangchunshu Zhou, Dong-Ho Lee, Ravi Kiran Selvam, Seyeon Lee 等ICLR 2021 · 被引用 73 次
- Self-Supervised Logic Induction for Explainable Fuzzy Temporal Commonsense ReasoningBibo Cai, Xiao Ding, Zhouhao Sun, Bing Qin 等AAAI 2023 · 被引用 11 次
- Pre-training Language Models with Deterministic Factual KnowledgeShaobo Li, Xiaoguang Li, Lifeng Shang, Chengjie Sun 等EMNLP 2022 · 被引用 13 次
- ERNIE 2.0: A Continual Pre-Training Framework for Language UnderstandingYu Sun, Shuohuan Wang, Yu-Kun Li, Shikun Feng 等AAAI 2020 · 被引用 885 次
- A Generative Approach for Script Event Prediction via Contrastive Fine-TuningFangqi Zhu, Jun Gao, Changlong Yu, Wei Wang 等AAAI 2023 · 被引用 22 次
