Warp to the Future: Joint Forecasting of Features and Feature Motion
Josip Saric, Marin Orsic, Tonci Antunovic, Sacha Vrazic, Sinisa Segvic
Abstract
We address anticipation of scene development by forecasting semantic segmentation of future frames. Several previous works approach this problem by F2F (featureto-feature) forecasting where future features are regressed from observed features. Different from previous work, we consider a novel F2M (feature-to-motion) formulation, which performs the forecast by warping observed features according to regressed feature flow. This formulation models a causal relationship between the past and the future, and regularizes inference by reducing dimensionality of the forecasting target. However, emergence of future scenery which was not visible in observed frames can not be explained by warping. We propose to address this issue by complementing F2M forecasting with the classic F2F approach. We realize this idea as a multi-head F2MF model built atop shared features. Experiments show that the F2M head prevails in static parts of the scene while the F2F head kicks-in to fill-in the novel regions. The proposed F2MF model operates in synergy with correlation features and outperforms all previous approaches both in short-term and mid-term forecast on the Cityscapes dataset.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers7
- Real-time Object Detection for Streaming PerceptionJinrong Yang, Songtao Liu, Zeming Li, Xiaoping Li et al.CVPR 2022 · 61 citations
- DINO-Foresight: Looking into the Future with DINOEfstathios Karypidis, Ioannis Kakogeorgiou, Spyridon Gidaris, Nikos KomodakisNeurIPS 2025 · 52 citations
- Predictive Feature Learning for Future Segmentation PredictionZihang Lin, Jiangxin Sun, Jianfang Hu, Qi-Zhi Yu et al.ICCV 2021 · 18 citations
- Long-term Video Frame Interpolation via Feature PropagationDawit Mureja Argaw, In So KweonCVPR 2022 · 12 citations
- Joint Forecasting of Panoptic Segmentations with Difference AttentionColin Graber, Cyril Jazra, Wenjie Luo, Liangyan Gui et al.CVPR 2022 · 3 citations
Builds on1
Related papers
- VFMF: Dense Forecasting by Generating Foundation Model FeaturesGabrijel Boduljak, Yushi Lan, Christian Rupprecht, Andrea VedaldiICML 2026
- Semantic Complete Scene Forecasting from a 4D Dynamic Point Cloud SequenceZifan Wang, Zhuorui Ye, Haoran Wu, Junyu Chen et al.AAAI 2024 · 8 citations
- Learning Semantic-Aware Dynamics for Video PredictionXinzhu Bei, Yanchao Yang, Stefano SoattoCVPR 2021
- Future Video Synthesis With Object Motion PredictionYue Wu, Rongrong Gao, Jaesik Park, Qifeng ChenCVPR 2020
- Multimodal Future Localization and Emergence Prediction for Objects in Egocentric View With a Reachability PriorOsama Makansi, Özgün Çiçek, Kevin Buchicchio, Thomas BroxCVPR 2020
