Self-Attention ConvLSTM for Spatiotemporal Prediction
Zhihui Lin, Maomao Li, Zhuobin Zheng, Yangyang Cheng, Chun Yuan
摘要
Spatiotemporal prediction is challenging due to the complex dynamic motion and appearance changes. Existing work concentrates on embedding additional cells into the standard ConvLSTM to memorize spatial appearances during the prediction. These models always rely on the convolution layers to capture the spatial dependence, which are local and inefficient. However, long-range spatial dependencies are significant for spatial applications. To extract spatial features with both global and local dependencies, we introduce the self-attention mechanism into ConvLSTM. Specifically, a novel self-attention memory (SAM) is proposed to memorize features with long-range dependencies in terms of spatial and temporal domains. Based on the self-attention, SAM can produce features by aggregating features across all positions of both the input itself and memory features with pair-wise similarity scores. Moreover, the additional memory is updated by a gating mechanism on aggregated features and an established highway with the memory of the previous time step. Therefore, through SAM, we can extract features with long-range spatiotemporal dependencies. Furthermore, we embed the SAM into a standard ConvLSTM to construct a self-attention ConvLSTM (SA-ConvLSTM) for the spatiotemporal prediction. In experiments, we apply the SA-ConvLSTM to perform frame prediction on the MovingMNIST and KTH datasets and traffic flow prediction on the TexiBJ dataset. Our SA-ConvLSTM achieves state-of-the-art results on both datasets with fewer parameters and higher time efficiency than previous state-of-the-art method.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper9
- PDFormer: Propagation Delay-Aware Dynamic Long-Range Transformer for Traffic Flow PredictionJiawei Jiang, Chengkai Han, Wayne Xin Zhao, Jingyuan WangAAAI 2023 · 被引用 542 次
- UniST: A Prompt-Empowered Universal Model for Urban Spatio-Temporal PredictionYuan Yuan, Jingtao Ding, Jie Feng, Depeng Jin 等KDD 2024 · 被引用 75 次
- Iso-Dream: Isolating and Leveraging Noncontrollable Visual Dynamics in World ModelsMinting Pan, Xiangming Zhu, Yunbo Wang, Xiaokang YangNeurIPS 2022 · 被引用 74 次
- PastNet: Introducing Physical Inductive Biases for Spatio-temporal Video PredictionHao Wu, Fan Xu, Chong Chen, Xian-Sheng Hua 等ACM MM 2024 · 被引用 36 次
- Diffusion Transformers as Open-World Spatiotemporal Foundation ModelsYuan Yuan, Chonghua Han, Jingtao Ding, Guozhen Zhang 等NeurIPS 2025 · 被引用 16 次
它引用的顶会 Paper1
相关 Paper
- SwinLSTM: Improving Spatiotemporal Prediction Accuracy using Swin Transformer and LSTMSong Tang, Chuang Li, Pu Zhang, Rongnian TangICCV 2023 · 被引用 117 次
- Modeling Citywide Crowd Flows using Attentive Convolutional LSTMChi Harold Liu, Chengzhe Piao, Xiaoxin Ma, Ye Yuan 等ICDE 2021 · 被引用 24 次
- ST-ABC: Spatio-Temporal Attention-Based Convolutional Network for Multi-Scale Lane-Level Traffic PredictionShuhao Li, Yue Cui, Libin Li, Weidong Yang 等ICDE 2024 · 被引用 13 次
- Convolutional State Space Models for Long-Range Spatiotemporal ModelingJimmy T. H. Smith, Shalini De Mello, Jan Kautz, Scott W. Linderman 等NeurIPS 2023 · 被引用 36 次
- SSL-STMFormer Self-Supervised Learning Spatio-Temporal Entanglement Transformer for Traffic Flow PredictionZetao Li, Zheng Hu, Peng Han, Yu Gu 等AAAI 2025 · 被引用 12 次
