PastNet: Introducing Physical Inductive Biases for Spatio-temporal Video Prediction
Hao Wu, Fan Xu, Chong Chen, Xian-Sheng Hua, Xiao Luo, Haixin Wang
摘要
In this paper, we investigate the challenge of spatio-temporal video prediction, which involves generating future videos based on historical data streams. Existing approaches typically utilize external information such as semantic maps to enhance video prediction, which often neglect the inherent physical knowledge embedded within videos. Furthermore, their high computational demands could impede their applications for high-resolution videos. To address these constraints, we introduce a novel approach called Physics-assisted Spatio-temporal Network (PastNet) for generating high-quality video prediction. The core of our PastNet lies in incorporating a spectral convolution operator in the Fourier domain, which efficiently introduces inductive biases from the underlying physical laws. Additionally, we employ a memory bank with the estimated intrinsic dimensionality to discretize local features during the processing of complex spatio-temporal signals, thereby reducing computational costs and facilitating efficient high-resolution video prediction. Extensive experiments on various widely-used datasets demonstrate the effectiveness and efficiency of the proposed PastNet compared with a range of state-of-the-art methods, particularly in high-resolution scenarios. Our code is available at https://github.com/easylearningscores/PastNet.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper24
- Revisiting Graph-Based Fraud Detection in Sight of Heterophily and SpectrumFan Xu, Nan Wang, Hao Wu, Xuezhi Wen 等AAAI 2024 · 被引用 72 次
- Earthfarsser: Versatile Spatio-Temporal Dynamical Systems Modeling in One ModelHao Wu, Yuxuan Liang, Wei Xiong, Zhengyang Zhou 等AAAI 2024 · 被引用 58 次
- NuwaDynamics: Discovering and Updating in Causal Spatio-Temporal ModelingKun Wang, Hao Wu, Yifan Duan, Guibin Zhang 等ICLR 2024 · 被引用 38 次
- Prometheus: Out-of-distribution Fluid Dynamics Modeling with Disentangled Graph ODEHao Wu, Huiyuan Wang, Kun Wang, Weiyan Wang 等ICML 2024 · 被引用 25 次
- PURE: Prompt Evolution with Graph ODE for Out-of-distribution Fluid Dynamics ModelingHao Wu, Changhu Wang, Fan Xu, Jinbao Xue 等NeurIPS 2024 · 被引用 23 次
它引用的顶会 Paper22
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- Fourier Neural Operator for Parametric Partial Differential EquationsZongyi Li, Nikola Borislavov Kovachki, Kamyar Azizzadenesheli, Burigede Liu 等ICLR 2021 · 被引用 3,911 次
- Vision Transformers for Dense PredictionRené Ranftl, Alexey Bochkovskiy, Vladlen KoltunICCV 2021 · 被引用 2,647 次
- DynamicViT: Efficient Vision Transformers with Dynamic Token SparsificationYongming Rao, Wenliang Zhao, Benlin Liu, Jiwen Lu 等NeurIPS 2021 · 被引用 1,343 次
- Self-Attention ConvLSTM for Spatiotemporal PredictionZhihui Lin, Maomao Li, Zhuobin Zheng, Yangyang Cheng 等AAAI 2020 · 被引用 347 次
相关 Paper
- Disentangling Physical Dynamics From Unknown Factors for Unsupervised Video PredictionVincent Le Guen, Nicolas ThomeCVPR 2020
- Enabling arbitrary inference in spatio-temporal dynamic systems: A physics-inspired perspectiveYan Ge, Zhengyang Zhou, Qihe Huang, Yuxuan Liang 等ICLR 2026
- ExtDM: Distribution Extrapolation Diffusion Model for Video PredictionZhicheng Zhang, Junyao Hu, Wentao Cheng, Danda Pani Paudel 等CVPR 2024 · 被引用 24 次
- STRPM: A Spatiotemporal Residual Predictive Model for High-Resolution Video PredictionZheng Chang, Xinfeng Zhang, Shanshe Wang, Siwei Ma 等CVPR 2022 · 被引用 57 次
- Ultrafast Video Attention Prediction with Coupled Knowledge DistillationKui Fu, Peipei Shi, Yafei Song, Shiming Ge 等AAAI 2020 · 被引用 11 次
