CamLiFlow: Bidirectional Camera-LiDAR Fusion for Joint Optical Flow and Scene Flow Estimation
Haisong Liu, Tao Lu, Yihui Xu, Jia Liu, Wenjie Li, Lijun Chen
Abstract
In this paper, we study the problem of jointly estimating the optical flow and scene flow from synchronized 2D and 3D data. Previous methods either employ a complex pipeline that splits the joint task into independent stages, or fuse 2D and 3D information in an "early-fusion" or "late-fusion" manner. Such one-size-fits-all approaches suffer from a dilemma of failing to fully utilize the characteristic of each modality or to maximize the inter-modality complementarity. To address the problem, we propose a novel end-to-end framework, which consists of 2D and 3D branches with multiple bidirectional fusion connections between them in specific layers. Different from previous work, we apply a point-based 3D branch to extract the LiDAR features, as it preserves the geometric structure of point clouds. To fuse dense image features and sparse point features, we propose a learnable operator named bidirectional camera-LiDAR fusion module (Bi-CLFM). We instantiate two types of the bidirectional fusion pipeline, one based on the pyramidal coarse-to-fine architecture (dubbed CamLiPWC), and the other one based on the recurrent all-pairs field transforms (dubbed CamLiRAFT). On FlyingThings3D, both CamLiPWC and CamLiRAFT surpass all existing methods and achieve up to a 47.9% reduction in 3D end-point-error from the best published result. Our best-performing model, CamLiRAFT, achieves an error of 4.26% on the KITTI Scene Flow benchmark, ranking 1st among all submissions with much fewer parameters. Besides, our methods have strong generalization performance and the ability to handle non-rigid motion. Code is available at https://github.com/MCG-NJU/CamLiFlow .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext e1e8460f-39a0-4dfa-a108-0e1e8e2c7a5fCited by top-tier papers25
- Unleash the Potential of Image Branch for Cross-modal 3D Object DetectionYifan Zhang, Qijian Zhang, Junhui Hou, Yixuan Yuan et al.NeurIPS 2023 · 36 citations
- RPEFlow: Multimodal Fusion of RGB-PointCloud-Event for Joint Optical Flow and Scene Flow EstimationZhexiong Wan, Yuxin Mao, Jing Zhang, Yuchao DaiICCV 2023 · 35 citations
- CVRecon: Rethinking 3D Geometric Feature Learning For Neural ReconstructionZiyue Feng, Liang Yang, Pengsheng Guo, Bing LiICCV 2023 · 28 citations
- GMSF: Global Matching Scene FlowYushan Zhang, Johan Edstedt, Bastian Wandt, Per-Erik Forssén et al.NeurIPS 2023 · 27 citations
- Density-invariant Features for Distant Point Cloud RegistrationQuan Liu, Hongzi Zhu, Yunsong Zhou, Hongyang Li et al.ICCV 2023 · 25 citations
Builds on18
- TransFusion: Robust LiDAR-Camera Fusion for 3D Object Detection with TransformersXuyang Bai, Zeyu Hu, Xinge Zhu, Qingqiu Huang et al.CVPR 2022 · 794 citations
- Pseudo-LiDAR++: Accurate Depth for 3D Object Detection in Autonomous DrivingYurong You, Yan Wang, Wei-Lun Chao, Divyansh Garg et al.ICLR 2020 · 439 citations
- RPVNet: A Deep and Efficient Range-Point-Voxel Fusion Network for LiDAR Point Cloud SegmentationJianyun Xu, Ruixiang Zhang, Jian Dou, Yushi Zhu et al.ICCV 2021 · 345 citations
- Neural-Guided RANSAC: Learning Where to Sample Model HypothesesEric Brachmann, Carsten RotherICCV 2019 · 282 citations
- MeteorNet: Deep Learning on Dynamic 3D Point Cloud SequencesXingyu Liu, Mengyuan Yan, Jeannette BohgICCV 2019 · 225 citations
Related papers
- RAFT-3D: Scene Flow Using Rigid-Motion EmbeddingsZachary Teed, Jia DengCVPR 2021
- PV-RAFT: Point-Voxel Correlation Fields for Scene Flow Estimation of Point CloudsYi Wei, Ziyi Wang, Yongming Rao, Jiwen Lu et al.CVPR 2021
- Feature-Level Collaboration: Joint Unsupervised Learning of Optical Flow, Stereo Depth and Camera MotionCheng Chi, Qingjie Wang, Tianyu Hao, Peng Guo et al.CVPR 2021
- Exploiting Rigidity Constraints for LiDAR Scene Flow EstimationGuanting Dong, Yueyi Zhang, Hanlin Li, Xiaoyan Sun et al.CVPR 2022 · 30 citations
- DELFlow: Dense Efficient Learning of Scene Flow for Large-Scale Point CloudsChensheng Peng, Guangming Wang, Xian Wan Lo, Xinrui Wu et al.ICCV 2023 · 19 citations
