Learning Semantic-Aware Dynamics for Video Prediction
Xinzhu Bei, Yanchao Yang, Stefano Soatto
摘要
We propose an architecture and training scheme to predict video frames by explicitly modeling dis-occlusions and capturing the evolution of semantically consistent regions in the video. The scene layout (semantic map) and motion (optical flow) are decomposed into layers, which are predicted and fused with their context to generate future layouts and motions. The appearance of the scene is warped from past frames using the predicted motion in co-visible regions; dis-occluded regions are synthesized with contentaware inpainting utilizing the predicted scene layout. The result is a predictive model that explicitly represents objects and learns their class-specific motion, which we evaluate on video prediction benchmarks.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper13
- Iso-Dream: Isolating and Leveraging Noncontrollable Visual Dynamics in World ModelsMinting Pan, Xiangming Zhu, Yunbo Wang, Xiaokang YangNeurIPS 2022 · 被引用 74 次
- Optimizing Video Prediction via Video Frame InterpolationYue Wu, Qiang Wen, Qifeng ChenCVPR 2022 · 被引用 47 次
- MMVP: Motion-Matrix-based Video PredictionYiqi Zhong, Luming Liang, Ilya Zharkov, Ulrich NeumannICCV 2023 · 被引用 39 次
- PastNet: Introducing Physical Inductive Biases for Spatio-temporal Video PredictionHao Wu, Fan Xu, Chong Chen, Xian-Sheng Hua 等ACM MM 2024 · 被引用 36 次
- STDiff: Spatio-Temporal Diffusion for Continuous Stochastic Video PredictionXi Ye, Guillaume-Alexandre BilodeauAAAI 2024 · 被引用 20 次
它引用的顶会 Paper3
- Disentangling Propagation and Generation for Video PredictionHang Gao, Huazhe Xu, Qi-Zhi Cai, Ruth Wang 等ICCV 2019 · 被引用 90 次
- Learning to Manipulate Individual Objects in an ImageYanchao Yang, Yutong Chen, Stefano SoattoCVPR 2020
- Future Video Synthesis With Object Motion PredictionYue Wu, Rongrong Gao, Jaesik Park, Qifeng ChenCVPR 2020
相关 Paper
- Every Frame Counts: Joint Learning of Video Segmentation and Optical FlowMingyu Ding, Zhe Wang, Bolei Zhou, Jianping Shi 等AAAI 2020 · 被引用 80 次
- WALDO: Future Video Synthesis using Object Layer Decomposition and Parametric Flow PredictionGuillaume Le Moing, Jean Ponce, Cordelia SchmidICCV 2023 · 被引用 7 次
- An Internal Learning Approach to Video InpaintingHaotian Zhang, Long Mai, Hailin Jin, Zhaowen Wang 等ICCV 2019 · 被引用 77 次
- Dynamic Shadow Unveils Invisible Semantics for Video OutpaintingRuilin Li, Hang Yu, Jiayan QiuNeurIPS 2025 · 被引用 2 次
- Flow-Guided Video Inpainting with Scene TemplatesDong Lao, Peihao Zhu, Peter Wonka, Ganesh SundaramoorthiICCV 2021 · 被引用 18 次
