SeMoLi: What Moves Together Belongs Together
Jenny Seidenschwarz, Aljosa Osep, Francesco Ferroni, Simon Lucey, Laura Leal-Taixé
Abstract
We tackle semi-supervised object detection based on motion cues. Recent results suggest that heuristic-based clustering methods in conjunction with object trackers can be used to pseudo-label instances of moving objects and use these as supervisory signals to train 3D object detectors in Lidar data without manual supervision. We re-think this approach and suggest that both, object detection, as well as motion-inspired pseudo-labeling, can be tackled in a data-driven manner. We leverage recent advances in scene flow estimation to obtain point trajectories from which we extract long-term, class-agnostic motion patterns. Revisiting correlation clustering in the context of message passing networks, we learn to group those motion patterns to cluster points to object instances. By estimating the full extent of the objects, we obtain per-scan 3D bounding boxes that we use to supervise a Lidar object detection network. Our method not only outperforms prior heuristic-based approaches (57.5 AP, + 14 improvement over prior work), more importantly, we show we can pseudo-label and train object detectors across datasets.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 4613a4d9-6e22-499e-85a8-8cd3da364480Cited by top-tier papers2
- MonoSOWA: Scalable Monocular 3D Object Detector Without Human AnnotationsJan Skvrna, Lukás NeumannICCV 2025 · 3 citations
- GOT-Edit: Geometry-Aware Generic Object Tracking via Online Model EditingShih-Fang Chen, Jun-Cheng Chen, I-Hong Jhuo, Yen-Yu LinICLR 2026 · 1 citation
Builds on9
- Deep Hough Voting for 3D Object Detection in Point CloudsCharles R. Qi, Or Litany, Kaiming He, Leonidas J. GuibasICCV 2019 · 1,467 citations
- Group-Free 3D Object Detection via TransformersZe Liu, Zheng Zhang, Yue Cao, Han Hu et al.ICCV 2021 · 368 citations
- Fast Neural Scene FlowXueqian Li, Jianqiao Zheng, Francesco Ferroni, Jhony Kaesemodel Pontes et al.ICCV 2023 · 41 citations
- Neural Prior for Trajectory EstimationChaoyang Wang, Xueqian Li, Jhony Kaesemodel Pontes, Simon LuceyCVPR 2022 · 19 citations
- RAMA: A Rapid Multicut Algorithm on GPUAhmed Abbas, Paul SwobodaCVPR 2022 · 7 citations
Related papers
- UNION: Unsupervised 3D Object Detection using Object Appearance-based Pseudo-ClassesTed de Vries Lentsch, Holger Caesar, Dariu GavrilaNeurIPS 2024 · 30 citations
- Seg2Box: 3D Object Detection by Point-Wise Semantics SupervisionMaoji Zheng, Ziyu Xu, Qiming Xia, Hai Wu et al.AAAI 2025 · 3 citations
- RigidFlow: Self-Supervised Scene Flow Learning on Point Clouds by Local Rigidity PriorRuibo Li, Chi Zhang, Guosheng Lin, Zhe Wang et al.CVPR 2022 · 47 citations
- MixSup: Mixed-grained Supervision for Label-efficient LiDAR-based 3D Object DetectionYuxue Yang, Lue Fan, Zhaoxiang ZhangICLR 2024 · 11 citations
- 3DSFLabelling: Boosting 3D Scene Flow Estimation by Pseudo Auto-LabellingChaokang Jiang, Guangming Wang, Jiuming Liu, Hesheng Wang et al.CVPR 2024
