Self-Attentive Hawkes Process
Qiang Zhang, Aldo Lipani, Ömer Kirnap, Emine Yilmaz
Abstract
Capturing the occurrence dynamics is crucial to predicting which type of events will happen next and when. A common method to do this is through Hawkes processes. To enhance their capacity, recurrent neural networks (RNNs) have been incorporated due to RNNs successes in processing sequential data such as languages. Recent evidence suggests that self-Attention is more competent than RNNs in dealing with languages. However, we are unaware of the effectiveness of self-Attention in the context of Hawkes processes. This study aims to fill the gap by designing a self-Attentive Hawkes process (SAHP). SAHP employs self-Attention to summarise the influence of history events and compute the probability of the next event. One deficit of the conventional selfattention, when applied to event sequences, is that its positional encoding only considers the order of a sequence ignoring the time intervals between events. To overcome this deficit, we modify its encoding by translating time intervals into phase shifts of sinusoidal functions. Experiments on goodness-of-fit and prediction tasks show the improved capability of SAHP. Furthermore, SAHP is more interpretable than RNN-based counterparts because the learnt attention weights reveal contributions of one event type to the happening of another type. To the best of our knowledge, this is the first work that studies the effectiveness of self-Attention in Hawkes processes.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers60
- ContiFormer: Continuous-Time Transformer for Irregular Time Series ModelingYuqi Chen, Kan Ren, Yansen Wang, Yuchen Fang et al.NeurIPS 2023 · 131 citations
- Transformer Embeddings of Irregularly Spaced Events and Their ParticipantsHongyuan Mei, Chenghao Yang, Jason EisnerICLR 2022 · 98 citations
- HYPRO: A Hybridly Normalized Probabilistic Model for Long-Horizon Prediction of Event SequencesSiqiao Xue, Xiaoming Shi, James Y. Zhang, Hongyuan MeiNeurIPS 2022 · 65 citations
- Mobility-LLM: Learning Visiting Intentions and Travel Preference from Human Mobility Data with Large Language ModelsLetian Gong, Yan Lin, Xinyue Zhang, Yiwen Lu et al.NeurIPS 2024 · 59 citations
- EasyTPP: Towards Open Benchmarking Temporal Point ProcessesSiqiao Xue, Xiaoming Shi, Zhixuan Chu, Yan Wang et al.ICLR 2024 · 53 citations
Builds on1
Related papers
- Transformer Hawkes ProcessSimiao Zuo, Haoming Jiang, Zichong Li, Tuo Zhao et al.ICML 2020 · 382 citations
- Dynamic Hawkes Processes for Discovering Time-evolving Communities' States behind Diffusion ProcessesMaya Okawa, Tomoharu Iwata, Yusuke Tanaka, Hiroyuki Toda et al.KDD 2021 · 8 citations
- TREND: TempoRal Event and Node Dynamics for Graph Representation LearningZhihao Wen, Yuan FangWWW 2022 · 113 citations
- Attentive Neural Point Processes for Event ForecastingYulong GuAAAI 2021 · 24 citations
- EIAN: Explicit Interaction-aware Attention Network for Interpretable Event ModelingJiping Zhang, Hua Zhu, Hong Huang, Yongkang Zhou et al.WWW 2026
