DiffSF: Diffusion Models for Scene Flow Estimation
Yushan Zhang, Bastian Wandt, Maria Magnusson, Michael Felsberg
Abstract
Scene flow estimation is an essential ingredient for a variety of real-world applications, especially for autonomous agents, such as self-driving cars and robots. While recent scene flow estimation approaches achieve a reasonable accuracy, their applicability to real-world systems additionally benefits from a reliability measure. Aiming at improving accuracy while additionally providing an estimate for uncertainty, we propose DiffSF that combines transformer-based scene flow estimation with denoising diffusion models. In the diffusion process, the ground truth scene flow vector field is gradually perturbed by adding Gaussian noise. In the reverse process, starting from randomly sampled Gaussian noise, the scene flow vector field prediction is recovered by conditioning on a source and a target point cloud. We show that the diffusion process greatly increases the robustness of predictions compared to prior approaches resulting in state-of-the-art performance on standard scene flow estimation benchmarks. Moreover, by sampling multiple times with different initial states, the denoising process predicts multiple hypotheses, which enables measuring the output uncertainty, allowing our approach to detect a majority of the inaccurate predictions. The code is available at https://github.com/ZhangYushan3/DiffSF.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 715dccc9-f9dc-48e5-98a6-dc3f5ae7fc7aCited by top-tier papers5
- DeltaFlow: An Efficient Multi-frame Scene Flow Estimation MethodQingwen Zhang, Xiaomeng Zhu, Yushan Zhang, Yixi Cai et al.NeurIPS 2025 · 8 citations
- TeFlow: Enabling Multi-frame Supervision for Self-Supervised Feed-forward Scene Flow EstimationQingwen Zhang, Chenhan Jiang, Xiaomeng Zhu, Yunqi Miao et al.CVPR 2026 · 5 citations
- DiffPCI: Large Motion Point Cloud Frame Interpolation with Diffusion ModelTianyu Zhang, Haobo Jiang, Jian Yang, Jin XieICCV 2025 · 1 citation
- GenFlow3D: Generative Scene Flow Estimation and Prediction on Point Cloud SequencesHanlin Li, Wenming Weng, Yueyi Zhang, Zhiwei XiongICCV 2025 · 1 citation
- U^2Flow: Uncertainty-Aware Unsupervised Optical Flow EstimationXunpei Sun, Wenwei Lin, Yi Chang, Gang ChenCVPR 2026
Builds on21
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 11,743 citations
- DiffusionDet: Diffusion Model for Object DetectionShoufa Chen, Peize Sun, Yibing Song, Ping LuoICCV 2023 · 715 citations
- Label-Efficient Semantic Segmentation with Diffusion ModelsDmitry Baranchuk, Andrey Voynov, Ivan Rubachev, Valentin Khrulkov et al.ICLR 2022 · 700 citations
Related papers
- DifFlow3D: Toward Robust Uncertainty-Aware Scene Flow Estimation with Iterative Diffusion-Based RefinementJiuming Liu, Guangming Wang, Weicai Ye, Chaokang Jiang et al.CVPR 2024
- GMSF: Global Matching Scene FlowYushan Zhang, Johan Edstedt, Bastian Wandt, Per-Erik Forssén et al.NeurIPS 2023 · 27 citations
- DiffusionSfM: Predicting Structure and Motion via Ray Origin and Endpoint DiffusionQitao Zhao, Amy Lin, Jeff Tan, Jason Y. Zhang et al.CVPR 2025
- SCTN: Sparse Convolution-Transformer Network for Scene Flow EstimationBing Li, Cheng Zheng, Silvio Giancola, Bernard GhanemAAAI 2022 · 50 citations
- ConsistentCity: Semantic Flow-Guided Occupancy DiT for Temporally Consistent Driving Scene SynthesisBenjin Zhu, Xiaogang Wang, Hongsheng LiICCV 2025 · 1 citation
