Decomposable Transformer Point Processes
Aristeidis Panos
Abstract
The standard paradigm of modeling marked point processes is by parameterizing the intensity function using an attention-based (Transformer-style) architecture. Despite the flexibility of these methods, their inference is based on the computationally intensive thinning algorithm. In this work, we propose a framework where the advantages of the attention-based architecture are maintained and the limitation of the thinning algorithm is circumvented. The framework depends on modeling the conditional distribution of inter-event times with a mixture of log-normals satisfying a Markov property and the conditional probability mass function for the marks with a Transformer-based architecture. The proposed method attains state-of-the-art performance in predicting the next event of a sequence given its history. The experiments also reveal the efficacy of the methods that do not rely on the thinning algorithm during inference over the ones they do. Finally, we test our method on the challenging long-horizon prediction task and find that it outperforms a baseline developed specifically for tackling this task; importantly, inference requires just a fraction of time compared to the thinning-based baseline.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext ef7fbb81-c04f-45aa-8024-e199935de8f7Cited by top-tier papers4
- In-Context Learning of Temporal Point Processes with Foundation Inference ModelsDavid Berghaus, Patrick Seifner, Kostadin Cvejoski, César Ali Ojeda Marin et al.ICLR 2026 · 8 citations
- TPP-SD: Accelerating Transformer Point Process Sampling with Speculative DecodingShukai Gong, Yiyang Fu, Fengyuan Ran, Quyu Kong et al.NeurIPS 2025 · 3 citations
- Addressing Mark Imbalance in Integration-free Marked Temporal Point ProcessesSishun Liu, Ke Deng, Yongli Ren, Yan Wang et al.NeurIPS 2025
- Long-range Modeling and Processing of Multimodal Event SequencesJichu Li, Yilun Zhong, Zhiting Li, Feng Zhou et al.ICLR 2026
Builds on7
- Transformer Hawkes ProcessSimiao Zuo, Haoming Jiang, Zichong Li, Tuo Zhao et al.ICML 2020 · 382 citations
- Self-Attentive Hawkes ProcessQiang Zhang, Aldo Lipani, Ömer Kirnap, Emine YilmazICML 2020 · 254 citations
- Intensity-Free Learning of Temporal Point ProcessesOleksandr Shchur, Marin Bilos, Stephan GünnemannICLR 2020 · 210 citations
- Transformer Embeddings of Irregularly Spaced Events and Their ParticipantsHongyuan Mei, Chenghao Yang, Jason EisnerICLR 2022 · 98 citations
- HYPRO: A Hybridly Normalized Probabilistic Model for Long-Horizon Prediction of Event SequencesSiqiao Xue, Xiaoming Shi, James Y. Zhang, Hongyuan MeiNeurIPS 2022 · 65 citations
Related papers
- Attentive Neural Point Processes for Event ForecastingYulong GuAAAI 2021 · 24 citations
- Transformers for Mixed-type Event SequencesFelix Draxler, Yang Meng, Kai Nelson, Lukas Laskowski et al.NeurIPS 2025 · 10 citations
- Conditional Generative Modeling for High-dimensional Marked Temporal Point ProcessesZheng Dong, Zekai Fan, Shixiang ZhuKDD 2025 · 1 citation
- ITPP: Learning Disentangled Event Dynamics in Marked Temporal Point ProcessesWang-Tao Zhou, Zhao Kang, Ke Yan, Ling TianAAAI 2026
- ProActive: Self-Attentive Temporal Point Process Flows for Activity SequencesVinayak Gupta, Srikanta BedathurKDD 2022 · 1 citation
