Disentangling Physical Dynamics From Unknown Factors for Unsupervised Video Prediction
Vincent Le Guen, Nicolas Thome
摘要
Leveraging physical knowledge described by partial differential equations (PDEs) is an appealing way to improve unsupervised video prediction methods. Since physics is too restrictive for describing the full visual content of generic videos, we introduce PhyDNet, a two-branch deep architecture, which explicitly disentangles PDE dynamics from unknown complementary information. A second contribution is to propose a new recurrent physical cell (PhyCell), inspired from data assimilation techniques, for performing PDE-constrained prediction in latent space. Extensive experiments conducted on four various datasets show the ability of PhyDNet to outperform state-of-the-art methods. Ablation studies also highlight the important gain brought out by both disentanglement and PDE-constrained prediction. Finally, we show that PhyDNet presents interesting features for dealing with missing data and long-term forecasting.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper72
- Earthformer: Exploring Space-Time Transformers for Earth System ForecastingZhihan Gao, Xingjian Shi, Hao Wang, Yi Zhu 等NeurIPS 2022 · 被引用 410 次
- SimVP: Simpler yet Better Video PredictionZhangyang Gao, Cheng Tan, Lirong Wu, Stan Z. LiCVPR 2022 · 被引用 313 次
- Cloth-Changing Person Re-identification from A Single Image with Gait Prediction and RegularizationXin Jin, Tianyu He, Kecheng Zheng, Zhiheng Yin 等CVPR 2022 · 被引用 174 次
- PreDiff: Precipitation Nowcasting with Latent Diffusion ModelsZhihan Gao, Xingjian Shi, Boran Han, Hao Wang 等NeurIPS 2023 · 被引用 171 次
- Stochastic Latent Residual Video PredictionJean-Yves Franceschi, Edouard Delasalles, Mickaël Chen, Sylvain Lamprier 等ICML 2020 · 被引用 166 次
它引用的顶会 Paper3
- Improved Conditional VRNNs for Video PredictionLluís Castrejón, Nicolas Ballas, Aaron C. CourvilleICCV 2019 · 被引用 177 次
- Disentangling Propagation and Generation for Video PredictionHang Gao, Huazhe Xu, Qi-Zhi Cai, Ruth Wang 等ICCV 2019 · 被引用 90 次
- Compositional Video PredictionYufei Ye, Maneesh Singh, Abhinav Gupta, Shubham TulsianiICCV 2019 · 被引用 84 次
相关 Paper
- PastNet: Introducing Physical Inductive Biases for Spatio-temporal Video PredictionHao Wu, Fan Xu, Chong Chen, Xian-Sheng Hua 等ACM MM 2024 · 被引用 36 次
- Learning Physics From Video: Unsupervised Physical Parameter Estimation for Continuous Dynamical SystemsAlejandro Castañeda Garcia, Jan Warchocki, Jan van Gemert, Daan Brinks 等CVPR 2025
- Disentangled Generative Models for Robust Prediction of System DynamicsStathi Fotiadis, Mario Lino Valencia, Shunlong Hu, Stef Garasto 等ICML 2023 · 被引用 14 次
- PDE-Driven Spatiotemporal DisentanglementJérémie Donà, Jean-Yves Franceschi, Sylvain Lamprier, Patrick GallinariICLR 2021 · 被引用 11 次
- MoAlign: Motion-Centric Representation Alignment for Video Diffusion ModelsAritra Bhowmik, Denis Korzhenkov, Cees G. M. Snoek, Amir Habibian 等ICLR 2026 · 被引用 15 次
