Continuous-Time Attention for Sequential Learning
Jen-Tzung Chien, Yi-Hsiang Chen
Abstract
Attention mechanism is crucial for sequential learning where a wide range of applications have been successfully developed. This mechanism is basically trained to spotlight on the region of interest in hidden states of sequence data. Most of the attention methods compute the attention score through relating between a query and a sequence where the discrete-time state trajectory is represented. Such a discrete-time attention could not directly attend the continuous-time trajectory which is represented via neural differential equation (NDE) combined with recurrent neural network. This paper presents a new continuous-time attention method for sequential learning which is tightly integrated with NDE to construct an attentive continuous-time state machine. The continuous-time attention is performed at all times over the hidden states for different kinds of irregular time signals. The missing information in sequence data due to sampling loss, especially in presence of long sequence, can be seamlessly compensated and attended in learning representation. The experiments on irregular sequence samples from human activities, dialogue sentences and medical features show the merits of the proposed continuous-time attention for activity recognition, sentiment classification and mortality prediction, respectively.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 9d822d18-45ae-4576-8f75-6dcdae7bb2e3Cited by top-tier papers2
- IVP-VAE: Modeling EHR Time Series with Initial Value Problem SolversJingge Xiao, Leonie Basso, Wolfgang Nejdl, Niloy Ganguly et al.AAAI 2024 · 14 citations
- HOP to the Next Tasks and Domains for Continual Learning in NLPUmberto Michieli, Mete OzayAAAI 2024 · 3 citations
Builds on2
Related papers
- DATA-GRU: Dual-Attention Time-Aware Gated Recurrent Unit for Irregular Multivariate Time SeriesQingxiong Tan, Mang Ye, Baoyao Yang, Siqi Liu et al.AAAI 2020 · 135 citations
- Continuous-Time Attention: PDE-Guided Mechanisms for Long-Sequence TransformersYukun Zhang, Xueqing ZhouEMNLP 2025
- Modeling Irregular Time Series with Continuous Recurrent UnitsMona Schirmer, Mazin Eltayeb, Stefan Lessmann, Maja RudolphICML 2022 · 135 citations
- ContiFormer: Continuous-Time Transformer for Irregular Time Series ModelingYuqi Chen, Kan Ren, Yansen Wang, Yuchen Fang et al.NeurIPS 2023 · 131 citations
- Efficient Anomaly Detection of Irregular Sequences in Ct-Echo Model SpaceAo Chen, Xiren Zhou, Huanhuan ChenAAAI 2025 · 2 citations
