Semantic Complete Scene Forecasting from a 4D Dynamic Point Cloud Sequence
Zifan Wang, Zhuorui Ye, Haoran Wu, Junyu Chen, Li Yi
Abstract
We study a new problem of semantic complete scene forecasting (SCSF) in this work. Given a 4D dynamic point cloud sequence, our goal is to forecast the complete scene corresponding to the future next frame along with its semantic labels. To tackle this challenging problem, we properly model the synergetic relationship between future forecasting and semantic scene completion through a novel network named SCSFNet. SCSFNet leverages a hybrid geometric representation for high-resolution complete scene forecasting. To leverage multi-frame observation as well as the understanding of scene dynamics to ease the completion task, SCSFNet introduces an attention-based skip connection scheme. To ease the need to model occlusion variations and to better focus on the occluded part, SCSFNet utilizes auxiliary visibility grids to guide the forecasting task. To evaluate the effectiveness of SCSFNet, we conduct experiments on various benchmarks including two large-scale indoor benchmarks we contributed and the outdoor SemanticKITTI benchmark. Extensive experiments show SCSFNet outperforms baseline methods on multiple metrics by a large margin, and also prove the synergy between future forecasting and semantic scene completion.The project page with code is available at scsfnet.github.io.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers2
- GaussianPrediction: Dynamic 3D Gaussian Prediction for Motion Extrapolation and Free View SynthesisBoming Zhao, Yuan Li, Ziyu Sun, Lin Zeng et al.SIGGRAPH 2024 · 17 citations
- ProxyTransformation: Preshaping Point Cloud Manifold With Proxy Attention For 3D Visual GroundingQihang Peng, Henry Zheng, Gao HuangCVPR 2025
Builds on9
- SemanticKITTI: A Dataset for Semantic Scene Understanding of LiDAR SequencesJens Behley, Martin Garbade, Andres Milioto, Jan Quenzel et al.ICCV 2019 · 2,345 citations
- Disentangling Propagation and Generation for Video PredictionHang Gao, Huazhe Xu, Qi-Zhi Cai, Ruth Wang et al.ICCV 2019 · 90 citations
- Cascaded Context Pyramid for Full-Resolution 3D Semantic Scene CompletionPingping Zhang, Wei Liu, Yinjie Lei, Huchuan Lu et al.ICCV 2019 · 79 citations
- ForkNet: Multi-Branch Volumetric Semantic Completion From a Single Depth ImageYida Wang, David Joseph Tan, Nassir Navab, Federico TombariICCV 2019 · 67 citations
- Point Cloud Forecasting as a Proxy for 4D Occupancy ForecastingTarasha Khurana, Peiyun Hu, David Held, Deva RamananCVPR 2023
Related papers
- Point Cloud Semantic Scene Completion from RGB-D ImagesShoulong Zhang, Shuai Li, Aimin Hao, Hong QinAAAI 2021 · 13 citations
- Point Cloud Completion by Skip-Attention Network With Hierarchical FoldingXin Wen, Tianyang Li, Zhizhong Han, Yu-Shen LiuCVPR 2020
- SDNet: LiDAR Semantic Scene Completion with Sparse-Dense Fusion and Input-Aware Label RefinementTingming Bai, Zhiyu Xiang, Peng Xu, Tianyu Pu et al.AAAI 2026
- Learning Temporal 3D Semantic Scene Completion via Optical Flow GuidanceMeng Wang, Fan Wu, Ruihui Li, Yunchuan Qin et al.NeurIPS 2025 · 4 citations
- CasFusionNet: A Cascaded Network for Point Cloud Semantic Scene Completion by Dense Feature FusionJinfeng Xu, Xianzhi Li, Yuan Tang, Qiao Yu et al.AAAI 2023 · 19 citations
