Mono-SF: Multi-View Geometry Meets Single-View Depth for Monocular Scene Flow Estimation of Dynamic Traffic Scenes
Fabian Brickwedde, Steffen Abraham, Rudolf Mester
Abstract
Existing 3D scene flow estimation methods provide the 3D geometry and 3D motion of a scene and gain a lot of interest, for example in the context of autonomous driving. These methods are traditionally based on a temporal series of stereo images. In this paper, we propose a novel monocular 3D scene flow estimation method, called Mono-SF. Mono-SF jointly estimates the 3D structure and motion of the scene by combining multi-view geometry and single-view depth information. Mono-SF considers that the scene flow should be consistent in terms of warping the reference image in the consecutive image based on the principles of multi-view geometry. For integrating single-view depth in a statistical manner, a convolutional neural network, called ProbDepthNet, is proposed. ProbDepthNet estimates pixel-wise depth distributions from a single image rather than single depth values. Additionally, as part of ProbDepthNet, a novel recalibration technique for regression problems is proposed to ensure well-calibrated distributions. Our experiments show that Mono-SF outperforms state-of-the-art monocular baselines and ablation studies support the Mono-SF approach and ProbDepthNet design.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext c2907c12-578e-433f-b299-0a6a554cd60fCited by top-tier papers18
- Neural Scene Flow PriorXueqian Li, Jhony Kaesemodel Pontes, Simon LuceyNeurIPS 2021 · 136 citations
- Fast Neural Scene FlowXueqian Li, Jianqiao Zheng, Francesco Ferroni, Jhony Kaesemodel Pontes et al.ICCV 2023 · 41 citations
- RPEFlow: Multimodal Fusion of RGB-PointCloud-Event for Joint Optical Flow and Scene Flow EstimationZhexiong Wan, Yuxin Mao, Jing Zhang, Yuchao DaiICCV 2023 · 35 citations
- Adaptive confidence thresholding for monocular depth estimationHyesong Choi, Hunsang Lee, Sunkyung Kim, Sunok Kim et al.ICCV 2021 · 31 citations
- Bring Event into RGB and LiDAR: Hierarchical Visual-Motion Fusion for Scene FlowHanyu Zhou, Yi Chang, Zhiwei ShiCVPR 2024 · 9 citations
Related papers
- Self-Supervised Monocular Scene Flow EstimationJunhwa Hur, Stefan RothCVPR 2020
- Multi-Object Discovery by Low-Dimensional Object MotionSadra Safadoust, Fatma GüneyICCV 2023 · 15 citations
- Consistent video depth estimationXuan Luo, Jia-Bin Huang, Richard Szeliski, Kevin Matzen et al.SIGGRAPH 2020 · 321 citations
- MonoMVSNet: Monocular Priors Guided Multi-View Stereo NetworkJianfei Jiang, Qiankun Liu, Haochen Yu, Hongyuan Liu et al.ICCV 2025 · 3 citations
- Scale-flow: Estimating 3D Motion from VideoHan Ling, Quansen Sun, Zhenwen Ren, Yazhou Liu et al.ACM MM 2022 · 7 citations
