Disentangling Physical Dynamics From Unknown Factors for Unsupervised Video Prediction
Vincent Le Guen, Nicolas Thome
Abstract
Leveraging physical knowledge described by partial differential equations (PDEs) is an appealing way to improve unsupervised video prediction methods. Since physics is too restrictive for describing the full visual content of generic videos, we introduce PhyDNet, a two-branch deep architecture, which explicitly disentangles PDE dynamics from unknown complementary information. A second contribution is to propose a new recurrent physical cell (PhyCell), inspired from data assimilation techniques, for performing PDE-constrained prediction in latent space. Extensive experiments conducted on four various datasets show the ability of PhyDNet to outperform state-of-the-art methods. Ablation studies also highlight the important gain brought out by both disentanglement and PDE-constrained prediction. Finally, we show that PhyDNet presents interesting features for dealing with missing data and long-term forecasting.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 39a84fa0-e096-41de-914d-9958cbf4f78cCited by top-tier papers72
- Earthformer: Exploring Space-Time Transformers for Earth System ForecastingZhihan Gao, Xingjian Shi, Hao Wang, Yi Zhu et al.NeurIPS 2022 · 410 citations
- SimVP: Simpler yet Better Video PredictionZhangyang Gao, Cheng Tan, Lirong Wu, Stan Z. LiCVPR 2022 · 313 citations
- Cloth-Changing Person Re-identification from A Single Image with Gait Prediction and RegularizationXin Jin, Tianyu He, Kecheng Zheng, Zhiheng Yin et al.CVPR 2022 · 174 citations
- PreDiff: Precipitation Nowcasting with Latent Diffusion ModelsZhihan Gao, Xingjian Shi, Boran Han, Hao Wang et al.NeurIPS 2023 · 171 citations
- Stochastic Latent Residual Video PredictionJean-Yves Franceschi, Edouard Delasalles, Mickaël Chen, Sylvain Lamprier et al.ICML 2020 · 166 citations
Builds on3
- Improved Conditional VRNNs for Video PredictionLluís Castrejón, Nicolas Ballas, Aaron C. CourvilleICCV 2019 · 177 citations
- Disentangling Propagation and Generation for Video PredictionHang Gao, Huazhe Xu, Qi-Zhi Cai, Ruth Wang et al.ICCV 2019 · 90 citations
- Compositional Video PredictionYufei Ye, Maneesh Singh, Abhinav Gupta, Shubham TulsianiICCV 2019 · 84 citations
Related papers
- PastNet: Introducing Physical Inductive Biases for Spatio-temporal Video PredictionHao Wu, Fan Xu, Chong Chen, Xian-Sheng Hua et al.ACM MM 2024 · 36 citations
- Learning Physics From Video: Unsupervised Physical Parameter Estimation for Continuous Dynamical SystemsAlejandro Castañeda Garcia, Jan Warchocki, Jan van Gemert, Daan Brinks et al.CVPR 2025
- Disentangled Generative Models for Robust Prediction of System DynamicsStathi Fotiadis, Mario Lino Valencia, Shunlong Hu, Stef Garasto et al.ICML 2023 · 14 citations
- PDE-Driven Spatiotemporal DisentanglementJérémie Donà, Jean-Yves Franceschi, Sylvain Lamprier, Patrick GallinariICLR 2021 · 11 citations
- MoAlign: Motion-Centric Representation Alignment for Video Diffusion ModelsAritra Bhowmik, Denis Korzhenkov, Cees G. M. Snoek, Amir Habibian et al.ICLR 2026 · 15 citations
