Video Prediction via Example Guidance
Jingwei Xu, Huazhe Xu, Bingbing Ni, Xiaokang Yang, Trevor Darrell
摘要
In video prediction tasks, one major challenge is to capture the multi-modal nature of future contents and dynamics. In this work, we propose a simple yet effective framework that can efficiently predict plausible future states. The key insight is that the potential distribution of a sequence could be approximated with analogous ones in a repertoire of training pool, namely, expert examples. By further incorporating a novel optimization scheme into the training procedure, plausible predictions can be sampled efficiently from distribution constructed from the retrieved examples. Meanwhile, our method could be seamlessly integrated with existing stochastic predictive models; significant enhancement is observed with comprehensive experiments in both quantitative and qualitative aspects. We also demonstrate the generalization ability to predict the motion of unseen class, i.e., without access to corresponding data during training phase. Project
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- Where are you heading? Dynamic Trajectory Prediction with Expert Goal ExamplesHe Zhao, Richard P. WildesICCV 2021 · 被引用 74 次
- STRPM: A Spatiotemporal Residual Predictive Model for High-Resolution Video PredictionZheng Chang, Xinfeng Zhang, Shanshe Wang, Siwei Ma 等CVPR 2022 · 被引用 57 次
- Multi-Objective Diverse Human Motion Prediction with Knowledge DistillationHengbo Ma, Jiachen Li, Ramtin Hosseini, Masayoshi Tomizuka 等CVPR 2022 · 被引用 41 次
- PastNet: Introducing Physical Inductive Biases for Spatio-temporal Video PredictionHao Wu, Fan Xu, Chong Chen, Xian-Sheng Hua 等ACM MM 2024 · 被引用 36 次
- A Frame is Worth One Token: Efficient Generative World Modeling with Delta TokensTommie Kerssies, Gabriele Berton, Ju He, Qihang Yu 等CVPR 2026 · 被引用 8 次
它引用的顶会 Paper4
- VideoFlow: A Conditional Flow-Based Model for Stochastic Video GenerationManoj Kumar, Mohammad Babaeizadeh, Dumitru Erhan, Chelsea Finn 等ICLR 2020 · 被引用 142 次
- Disentangling Propagation and Generation for Video PredictionHang Gao, Huazhe Xu, Qi-Zhi Cai, Ruth Wang 等ICCV 2019 · 被引用 90 次
- Compositional Video PredictionYufei Ye, Maneesh Singh, Abhinav Gupta, Shubham TulsianiICCV 2019 · 被引用 84 次
- Deep Kinematics Analysis for Monocular 3D Human Pose EstimationJingwei Xu, Zhenbo Yu, Bingbing Ni, Jiancheng Yang 等CVPR 2020
相关 Paper
- PlaySlot: Learning Inverse Latent Dynamics for Controllable Object-Centric Video Prediction and PlanningAngel Villar-Corrales, Sven BehnkeICML 2025
- Contextually Plausible and Diverse 3D Human Motion PredictionSadegh Aliakbarian, Fatemeh Sadat Saleh, Lars Petersson, Stephen Gould 等ICCV 2021 · 被引用 44 次
- Diverse Video Generation using a Gaussian Process TriggerGaurav Shrivastava, Abhinav ShrivastavaICLR 2021 · 被引用 22 次
- Optimizing Video Prediction via Video Frame InterpolationYue Wu, Qiang Wen, Qifeng ChenCVPR 2022 · 被引用 47 次
- Forecasting Characteristic 3D Poses of Human ActionsChristian Diller, Thomas A. Funkhouser, Angela DaiCVPR 2022 · 被引用 22 次
