PastNet: Introducing Physical Inductive Biases for Spatio-temporal Video Prediction
Hao Wu, Fan Xu, Chong Chen, Xian-Sheng Hua, Xiao Luo, Haixin Wang
Abstract
In this paper, we investigate the challenge of spatio-temporal video prediction, which involves generating future videos based on historical data streams. Existing approaches typically utilize external information such as semantic maps to enhance video prediction, which often neglect the inherent physical knowledge embedded within videos. Furthermore, their high computational demands could impede their applications for high-resolution videos. To address these constraints, we introduce a novel approach called Physics-assisted Spatio-temporal Network (PastNet) for generating high-quality video prediction. The core of our PastNet lies in incorporating a spectral convolution operator in the Fourier domain, which efficiently introduces inductive biases from the underlying physical laws. Additionally, we employ a memory bank with the estimated intrinsic dimensionality to discretize local features during the processing of complex spatio-temporal signals, thereby reducing computational costs and facilitating efficient high-resolution video prediction. Extensive experiments on various widely-used datasets demonstrate the effectiveness and efficiency of the proposed PastNet compared with a range of state-of-the-art methods, particularly in high-resolution scenarios. Our code is available at https://github.com/easylearningscores/PastNet.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 956f6fc2-aef7-440d-ac8c-832da5bdc410Cited by top-tier papers24
- Revisiting Graph-Based Fraud Detection in Sight of Heterophily and SpectrumFan Xu, Nan Wang, Hao Wu, Xuezhi Wen et al.AAAI 2024 · 72 citations
- Earthfarsser: Versatile Spatio-Temporal Dynamical Systems Modeling in One ModelHao Wu, Yuxuan Liang, Wei Xiong, Zhengyang Zhou et al.AAAI 2024 · 58 citations
- NuwaDynamics: Discovering and Updating in Causal Spatio-Temporal ModelingKun Wang, Hao Wu, Yifan Duan, Guibin Zhang et al.ICLR 2024 · 38 citations
- Prometheus: Out-of-distribution Fluid Dynamics Modeling with Disentangled Graph ODEHao Wu, Huiyuan Wang, Kun Wang, Weiyan Wang et al.ICML 2024 · 25 citations
- PURE: Prompt Evolution with Graph ODE for Out-of-distribution Fluid Dynamics ModelingHao Wu, Changhu Wang, Fan Xu, Jinbao Xue et al.NeurIPS 2024 · 23 citations
Builds on22
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Fourier Neural Operator for Parametric Partial Differential EquationsZongyi Li, Nikola Borislavov Kovachki, Kamyar Azizzadenesheli, Burigede Liu et al.ICLR 2021 · 3,911 citations
- Vision Transformers for Dense PredictionRené Ranftl, Alexey Bochkovskiy, Vladlen KoltunICCV 2021 · 2,647 citations
- DynamicViT: Efficient Vision Transformers with Dynamic Token SparsificationYongming Rao, Wenliang Zhao, Benlin Liu, Jiwen Lu et al.NeurIPS 2021 · 1,343 citations
- Self-Attention ConvLSTM for Spatiotemporal PredictionZhihui Lin, Maomao Li, Zhuobin Zheng, Yangyang Cheng et al.AAAI 2020 · 347 citations
Related papers
- Disentangling Physical Dynamics From Unknown Factors for Unsupervised Video PredictionVincent Le Guen, Nicolas ThomeCVPR 2020
- Enabling arbitrary inference in spatio-temporal dynamic systems: A physics-inspired perspectiveYan Ge, Zhengyang Zhou, Qihe Huang, Yuxuan Liang et al.ICLR 2026
- ExtDM: Distribution Extrapolation Diffusion Model for Video PredictionZhicheng Zhang, Junyao Hu, Wentao Cheng, Danda Pani Paudel et al.CVPR 2024 · 24 citations
- STRPM: A Spatiotemporal Residual Predictive Model for High-Resolution Video PredictionZheng Chang, Xinfeng Zhang, Shanshe Wang, Siwei Ma et al.CVPR 2022 · 57 citations
- Ultrafast Video Attention Prediction with Coupled Knowledge DistillationKui Fu, Peipei Shi, Yafei Song, Shiming Ge et al.AAAI 2020 · 11 citations
