BiTimeBERT: Extending Pre-Trained Language Representations with Bi-Temporal Information
Jiexin Wang, Adam Jatowt, Masatoshi Yoshikawa, Yi Cai
摘要
Time is an important aspect of documents and is used in a range of NLP and IR tasks. In this work, we investigate methods for incorporating temporal information during pre-training to further improve the performance on time-related tasks. Compared with common pre-trained language models like BERT which utilize synchronic document collections (e.g., BookCorpus and Wikipedia) as the training corpora, we use long-span temporal news article collection for building word representations. We introduce BiTimeBERT, a novel language representation model trained on a temporal collection of news articles via two new pre-training tasks, which harnesses two distinct temporal signals to construct time-aware language representations. The experimental results show that BiTimeBERT consistently outperforms BERT and other existing pre-trained models with substantial gains on different downstream NLP tasks and applications for which time is of importance (e.g., the accuracy improvement over BERT is 155% on the event time estimation task). 1
• Information systems → Content analysis and feature selection.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper8
- Do Language Models Have a Common Sense regarding Time? Revisiting Temporal Commonsense Reasoning in the Era of Large Language ModelsRaghav Jain, Daivik Sojitra, Arkadeep Acharya, Sriparna Saha 等EMNLP 2023 · 被引用 17 次
- Few-Shot Joint Multimodal Entity-Relation Extraction via Knowledge-Enhanced Cross-modal Prompt ModelLi Yuan, Yi Cai, Junsheng HuangACM MM 2024 · 被引用 9 次
- It's High Time: A Survey of Temporal Question AnsweringBhawna Piryani, Abdelrahman Abdallah, Jamshid Mozafari, Avishek Anand 等ACL 2026 · 被引用 6 次
- ComplexTempQA: A 100m Dataset for Complex Temporal Question AnsweringRaphael Gruber, Abdelrahman Abdallah, Michael Färber, Adam JatowtEMNLP 2025 · 被引用 2 次
- TempoFormer: A Transformer for Temporally-aware Representations in Change DetectionTalia Tseriotou, Adam Tsakalidis, Maria LiakataEMNLP 2024 · 被引用 2 次
它引用的顶会 Paper8
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah 等NeurIPS 2020 · 被引用 64,255 次
- Retrieval Augmented Language Model Pre-TrainingKelvin Guu, Kenton Lee, Zora Tung, Panupong Pasupat 等ICML 2020 · 被引用 2,937 次
- Pretrained Encyclopedia: Weakly Supervised Knowledge-Pretrained Language ModelWenhan Xiong, Jingfei Du, William Yang Wang, Veselin StoyanovICLR 2020 · 被引用 215 次
- SentiLARE: Sentiment-Aware Language Representation Learning with Linguistic KnowledgePei Ke, Haozhe Ji, Siyang Liu, Xiaoyan Zhu 等EMNLP 2020 · 被引用 138 次
- Analysing Lexical Semantic Change with Contextualised Word RepresentationsMario Giulianelli, Marco Del Tredici, Raquel FernándezACL 2020 · 被引用 118 次
相关 Paper
- Temporal Common Sense Acquisition with Minimal SupervisionBen Zhou, Qiang Ning, Daniel Khashabi, Dan RothACL 2020 · 被引用 76 次
- TimesBERT: A BERT-Style Foundation Model for Time Series UnderstandingHaoran Zhang, Yong Liu, Yunzhong Qiu, Haixuan Liu 等ACM MM 2025 · 被引用 7 次
- LinkBERT: Pretraining Language Models with Document LinksMichihiro Yasunaga, Jure Leskovec, Percy LiangACL 2022 · 被引用 463 次
- Modeling Document-Level Context for Event Detection via Important Context SelectionAmir Pouran Ben Veyseh, Minh Van Nguyen, Nghia Trung Ngo, Bonan Min 等EMNLP 2021 · 被引用 25 次
- Universal Sentence Representation Learning with Conditional Masked Language ModelZiyi Yang, Yinfei Yang, Daniel Cer, Jax Law 等EMNLP 2021 · 被引用 37 次
