Warp to the Future: Joint Forecasting of Features and Feature Motion
Josip Saric, Marin Orsic, Tonci Antunovic, Sacha Vrazic, Sinisa Segvic
摘要
We address anticipation of scene development by forecasting semantic segmentation of future frames. Several previous works approach this problem by F2F (featureto-feature) forecasting where future features are regressed from observed features. Different from previous work, we consider a novel F2M (feature-to-motion) formulation, which performs the forecast by warping observed features according to regressed feature flow. This formulation models a causal relationship between the past and the future, and regularizes inference by reducing dimensionality of the forecasting target. However, emergence of future scenery which was not visible in observed frames can not be explained by warping. We propose to address this issue by complementing F2M forecasting with the classic F2F approach. We realize this idea as a multi-head F2MF model built atop shared features. Experiments show that the F2M head prevails in static parts of the scene while the F2F head kicks-in to fill-in the novel regions. The proposed F2MF model operates in synergy with correlation features and outperforms all previous approaches both in short-term and mid-term forecast on the Cityscapes dataset.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper7
- Real-time Object Detection for Streaming PerceptionJinrong Yang, Songtao Liu, Zeming Li, Xiaoping Li 等CVPR 2022 · 被引用 61 次
- DINO-Foresight: Looking into the Future with DINOEfstathios Karypidis, Ioannis Kakogeorgiou, Spyridon Gidaris, Nikos KomodakisNeurIPS 2025 · 被引用 52 次
- Predictive Feature Learning for Future Segmentation PredictionZihang Lin, Jiangxin Sun, Jianfang Hu, Qi-Zhi Yu 等ICCV 2021 · 被引用 18 次
- Long-term Video Frame Interpolation via Feature PropagationDawit Mureja Argaw, In So KweonCVPR 2022 · 被引用 12 次
- Joint Forecasting of Panoptic Segmentations with Difference AttentionColin Graber, Cyril Jazra, Wenjie Luo, Liangyan Gui 等CVPR 2022 · 被引用 3 次
它引用的顶会 Paper1
相关 Paper
- VFMF: Dense Forecasting by Generating Foundation Model FeaturesGabrijel Boduljak, Yushi Lan, Christian Rupprecht, Andrea VedaldiICML 2026
- Semantic Complete Scene Forecasting from a 4D Dynamic Point Cloud SequenceZifan Wang, Zhuorui Ye, Haoran Wu, Junyu Chen 等AAAI 2024 · 被引用 8 次
- Learning Semantic-Aware Dynamics for Video PredictionXinzhu Bei, Yanchao Yang, Stefano SoattoCVPR 2021
- Future Video Synthesis With Object Motion PredictionYue Wu, Rongrong Gao, Jaesik Park, Qifeng ChenCVPR 2020
- Multimodal Future Localization and Emergence Prediction for Objects in Egocentric View With a Reachability PriorOsama Makansi, Özgün Çiçek, Kevin Buchicchio, Thomas BroxCVPR 2020
