Learning To Segment Rigid Motions From Two Frames
Gengshan Yang, Deva Ramanan
Abstract
Reference frame (b) Geometric (black: rigid background) (d) Our rigid motion (projected to 2D) (a) PointRend (trained on MSCOCO) (e) Our two-frame reconstruction Figure 1: (a) Many data-driven segmentation methods heavily rely on appearance cues, and fail for novel test scenes. For instance, PointRend [25] trained on MSCOCO fails to detect coral reef fishes even with a low confidence threshold of 0.1. (b) On the other hand, geometric motion segmentation [5, 58] generalizes to novel appearance, but fails due to noisy flow inputs and degenerate motion configurations. (c)-(e) We propose a neural architecture powered by geometric reasoning that decomposes a scene into a rigid background and multiple moving rigid bodies, parameterized by 3D rigid transformations. It demonstrates generalization ability to novel scenes and robustness to noisy inputs as well as motion degeneracies. The inferred rigid motions significantly improve depth and scene flow accuracy.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers14
- CRAFT: Cross-Attentional Flow Transformer for Robust Optical FlowXiuchao Sui, Shaohua Li, Xue Geng, Yan Wu et al.CVPR 2022 · 114 citations
- TAPIP3D: Tracking Any Point in Persistent 3D GeometryBowei Zhang, Lei Ke, Adam W. Harley, Katerina FragkiadakiNeurIPS 2025 · 79 citations
- CamLiFlow: Bidirectional Camera-LiDAR Fusion for Joint Optical Flow and Scene Flow EstimationHaisong Liu, Tao Lu, Yihui Xu, Jia Liu et al.CVPR 2022 · 64 citations
- Dynamo-Depth: Fixing Unsupervised Depth Estimation for Dynamical ScenesYihong Sun, Bharath HariharanNeurIPS 2023 · 58 citations
- GMSF: Global Matching Scene FlowYushan Zhang, Johan Edstedt, Bastian Wandt, Per-Erik Forssén et al.NeurIPS 2023 · 27 citations
Builds on7
- Neural-Guided RANSAC: Learning Where to Sample Model HypothesesEric Brachmann, Carsten RotherICCV 2019 · 282 citations
- Motion-Attentive Transition for Zero-Shot Video Object SegmentationTianfei Zhou, Shunzhou Wang, Yi Zhou, Yazhou Yao et al.AAAI 2020 · 210 citations
- Anchor Diffusion for Unsupervised Video Object SegmentationZhao Yang, Qiang Wang, Luca Bertinetto, Song Bai et al.ICCV 2019 · 127 citations
- Mono-SF: Multi-View Geometry Meets Single-View Depth for Monocular Scene Flow Estimation of Dynamic Traffic ScenesFabian Brickwedde, Steffen Abraham, Rudolf MesterICCV 2019 · 55 citations
- Self-Supervised Monocular Scene Flow EstimationJunhwa Hur, Stefan RothCVPR 2020
Related papers
- Weakly Supervised Learning of Rigid 3D Scene FlowZan Gojcic, Or Litany, Andreas Wieser, Leonidas J. Guibas et al.CVPR 2021
- TRACE: Learning 3D Gaussian Physical Dynamics from Multi-View VideosJinxi Li, Ziyang Song, Bo YangICCV 2025 · 3 citations
- GenMatter: Perceiving Physical Objects with Generative Matter ModelsEric Li, Arijit Dasgupta, Yoni Friedman, Mathieu Huot et al.CVPR 2026
- Unsupervised Multi-Object Segmentation by Predicting Probable Motion PatternsLaurynas Karazija, Subhabrata Choudhury, Iro Laina, Christian Rupprecht et al.NeurIPS 2022 · 24 citations
- Multi-body SE(3) Equivariance for Unsupervised Rigid Segmentation and Motion EstimationJia-Xing Zhong, Ta Ying Cheng, Yuhang He, Kai Lu et al.NeurIPS 2023 · 9 citations
