Transformers for Mixed-type Event Sequences
Felix Draxler, Yang Meng, Kai Nelson, Lukas Laskowski, Yibo Yang, Theofanis Karaletsos, Stephan Mandt
摘要
Event sequences appear widely in domains such as medicine, finance, and remote sensing, yet modeling them is challenging due to their heterogeneity: sequences often contain multiple event types with diverse structures—for example, electronic health records that mix discrete events like medical procedures with continuous lab measurements. Existing approaches either tokenize all entries, violating natural inductive biases, or ignore parts of the data to enforce a consistent structure. In this work, we propose a simple yet powerful Marked Temporal Point Process (MTPP) framework for modeling event sequences with flexible structure, using a single unified model. Our approach employs a single autoregressive transformer with discrete and continuous prediction heads, capable of modeling variable-length, mixed-type event sequences. The continuous head leverages an expressive normalizing flow to model continuous event attributes, avoiding the numerical integration required for inter-event times in most competing methods. Empirically, our model excels on both discrete-only and mixed-type sequences, improving prediction quality and enabling interpretable uncertainty quantification. We make our code public at https://github.com/czi-ai/FlexTPP .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- Deep Continuous-Time State-Space Models for Marked Event SequencesYuxin Chang, Alex Boyd, Cao (Danica) Xiao, Taha A. Kass-Hout 等NeurIPS 2025 · 被引用 11 次
- In-Context Learning of Temporal Point Processes with Foundation Inference ModelsDavid Berghaus, Patrick Seifner, Kostadin Cvejoski, César Ali Ojeda Marin 等ICLR 2026 · 被引用 8 次
- Parallel Token Prediction for Language ModelsFelix Draxler, Justus C. Will, Farrin Marouf Sofian, Theofanis Karaletsos 等ICLR 2026 · 被引用 6 次
- Net-Ev2: A Generative Simulator for Network Event EvolutionGuangyu Wang, Zhaonan WangKDD 2026 · 被引用 1 次
- Skipping the Zeros in Diffusion Models for Sparse Data GenerationPhil Sidney Ostheimer, Mayank Kumar Nagda, Andriy Balinskyy, Gabriel Rodrigues 等ICML 2026
它引用的顶会 Paper17
- Flamingo: a Visual Language Model for Few-Shot LearningJean-Baptiste Alayrac, Jeff Donahue, Pauline Luc, Antoine Miech 等NeurIPS 2022 · 被引用 6,707 次
- Efficiently Modeling Long Sequences with Structured State SpacesAlbert Gu, Karan Goel, Christopher RéICLR 2022 · 被引用 3,482 次
- Transformers are RNNs: Fast Autoregressive Transformers with Linear AttentionAngelos Katharopoulos, Apoorv Vyas, Nikolaos Pappas, François FleuretICML 2020 · 被引用 2,665 次
- Anomaly Detection in Time Series: A Comprehensive EvaluationSebastian Schmidl, Phillip Wenig, Thorsten PapenbrockVLDB 2022 · 被引用 578 次
- Transformer Hawkes ProcessSimiao Zuo, Haoming Jiang, Zichong Li, Tuo Zhao 等ICML 2020 · 被引用 382 次
相关 Paper
- ProActive: Self-Attentive Temporal Point Process Flows for Activity SequencesVinayak Gupta, Srikanta BedathurKDD 2022 · 被引用 1 次
- ITPP: Learning Disentangled Event Dynamics in Marked Temporal Point ProcessesWang-Tao Zhou, Zhao Kang, Ke Yan, Ling TianAAAI 2026
- Add and Thin: Diffusion for Temporal Point ProcessesDavid Lüdke, Marin Bilos, Oleksandr Shchur, Marten Lienen 等NeurIPS 2023 · 被引用 34 次
- Fast and Flexible Temporal Point Processes with Triangular MapsOleksandr Shchur, Nicholas Gao, Marin Bilos, Stephan GünnemannNeurIPS 2020 · 被引用 43 次
- Intensity-Free Learning of Temporal Point ProcessesOleksandr Shchur, Marin Bilos, Stephan GünnemannICLR 2020 · 被引用 210 次
