Pairwise Causality Guided Transformers for Event Sequences
Xiao Shou, Debarun Bhattacharjya, Tian Gao, Dharmashankar Subramanian, Oktie Hassanzadeh, Kristin P. Bennett
摘要
Although pairwise causal relations have been extensively studied in observational longitudinal analyses across many disciplines, incorporating knowledge of causal pairs into deep learning models for temporal event sequences remains largely unexplored. In this paper, we propose a novel approach for enhancing the performance of transformer-based models in multivariate event sequences by injecting pairwise qualitative causal knowledge such as ‘event Z amplifies future occurrences of event Y’. We establish a new framework for causal inference in temporal event sequences using a transformer architecture, providing a theoretical justification for our approach, and show how to obtain unbiased estimates of the proposed measure. Experimental results demonstrate that our approach outperforms several state-of-the-art models in terms of prediction accuracy by effectively leveraging knowledge about causal pairs. We also consider a unique application where we extract knowledge around sequences of societal events by generating them from a large language model, and demonstrate how a causal knowledge graph can help with event prediction in such sequences. Overall, our framework offers a practical means of improving the performance of transformer-based models in multivariate event sequences by explicitly exploiting pairwise causal information.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper13
- Chain-of-Thought Prompting Elicits Reasoning in Large Language ModelsJason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma 等NeurIPS 2022 · 被引用 22,562 次
- Locating and Editing Factual Associations in GPTKevin Meng, David Bau, Alex Andonian, Yonatan BelinkovNeurIPS 2022 · 被引用 3,415 次
- Estimating counterfactual treatment outcomes over time through adversarially balanced representationsIoana Bica, Ahmed M. Alaa, James Jordon, Mihaela van der SchaarICLR 2020 · 被引用 224 次
- Causal Transformer for Estimating Counterfactual OutcomesValentyn Melnychuk, Dennis Frauen, Stefan FeuerriegelICML 2022 · 被引用 146 次
- CAUSE: Learning Granger Causality from Event Sequences using Attribution MethodsWei Zhang, Thomas Kobber Panum, Somesh Jha, Prasad Chalasani 等ICML 2020 · 被引用 64 次
相关 Paper
- Transformer Hawkes ProcessSimiao Zuo, Haoming Jiang, Zichong Li, Tuo Zhao 等ICML 2020 · 被引用 382 次
- GraFT: Infusing Pre-trained Transformers with Relational Structure for Time Series ForecastingYuqi Yuan, Xiong Luo, Qiaojuan Peng, Wenbing ZhaoAAAI 2026
- Causal Graph based Event Reasoning using Semantic Relation ExpertsMahnaz Koupaee, Xueying Bai, Mudan Chen, Greg Durrett 等ACL 2025
- Causal Interpretation of Self-Attention in Pre-Trained TransformersRaanan Y. Rohekar, Yaniv Gurwicz, Shami NisimovNeurIPS 2023 · 被引用 62 次
- Robust Event Forecasting with Spatiotemporal Confounder LearningSonggaojun Deng, Huzefa Rangwala, Yue NingKDD 2022 · 被引用 9 次
