FS-I2P: A Hierarchical Focus–Sweep Registration Network with Dynamically Allocated Depth
Zhixin Cheng, Yujia Chen, Xujing Tao, Bohao Liao, Xiaotian Yin, Baoqun Yin, Tianzhu Zhang
Abstract
Image-to-point cloud registration is often challenged by viewpoint changes, cross-modal discrepancies, and repetitive textures, which induce scale ambiguity and consequently lead to erroneous correspondences. Recent detection-free methods alleviate this issue by leveraging multi-scale features and transformer-based interactions. However, they still suffer from attention drift across layers and intra-scale inconsistencies, hindering precise registration. Inspired by complex scene observation, we propose a ``Focus--Sweep'' paradigm and develop a Hierarchical Mamba Interaction Module within an SSM-based framework to enhance multi-level cross-modal feature association. In addition, we introduce a Dynamic Layer Allocation Strategy that adaptively determines the iteration depth to better exploit geometric constraints and improve matching robustness. Extensive experiments and ablations on two benchmarks, RGB-D Scenes V2 and 7-Scenes, demonstrate that our approach achieves state-of-the-art performance.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 29d7d9fb-6dcb-4123-aeb9-4a64ad9a5c3cCited by top-tier papers3
- GeoGuide: Hierarchical Geometric Guidance for Open-Vocabulary 3D Semantic SegmentationXujing Tao, Chuxin Wang, Yubo Ai, Zhixin Cheng et al.CVPR 2026 · 3 citations
- Rethinking 2D-3D Registration: A Novel Network for High-Value Zone Selection and Representation Consistency AlignmentZhixin Cheng, Bohao Liao, Jiacheng Deng, Xiaotian Yin et al.CVPR 2026 · 2 citations
- Beyond Logits: Coherent Hallucination Mitigation via Attention Contrastive DecodingYujia Chen, Rui Sun, Huayu Mai, Wangkai Li et al.ICML 2026
Builds on31
- Fully Convolutional Geometric FeaturesChristopher B. Choy, Jaesik Park, Vladlen KoltunICCV 2019 · 807 citations
- Geometric Transformer for Fast and Robust Point Cloud RegistrationZheng Qin, Hao Yu, Changjian Wang, Yulan Guo et al.CVPR 2022 · 436 citations
- Efficient LoFTR: Semi-Dense Local Feature Matching with Sparse-Like SpeedYifan Wang, Xingyi He, Sida Peng, Dongli Tan et al.CVPR 2024 · 126 citations
- Revisiting Domain Generalized Stereo Matching Networks from a Feature Consistency PerspectiveJiawei Zhang, Xiang Wang, Xiao Bai, Chen Wang et al.CVPR 2022 · 81 citations
- P2-Net: Joint Description and Detection of Local Features for Pixel and Point MatchingBing Wang, Changhao Chen, Zhaopeng Cui, Jie Qin et al.ICCV 2021 · 75 citations
Related papers
- Adaptive Agent Selection and Interaction Network for Image-to-Point Cloud RegistrationZhixin Cheng, Xiaotian Yin, Jiacheng Deng, Bohao Liao et al.AAAI 2026 · 4 citations
- 2D3D-MATR: 2D-3D Matching Transformer for Detection-free Registration between Images and Point CloudsMinhao Li, Zheng Qin, Zhirui Gao, Renjiao Yi et al.ICCV 2023 · 30 citations
- CA-I2P: Channel-Adaptive Registration Network with Global Optimal SelectionZhixin Cheng, Jiacheng Deng, Xinjun Li, Xiaotian Yin et al.ICCV 2025 · 4 citations
- 3DET-Mamba: Causal Sequence Modelling for End-to-End 3D Object DetectionMingsheng Li, Jiakang Yuan, Sijin Chen, Lin Zhang et al.NeurIPS 2024 · 5 citations
- Bridge 2D-3D: Uncertainty-aware Hierarchical Registration Network with Domain AlignmentZhixin Cheng, Jiacheng Deng, Xinjun Li, Baoqun Yin et al.AAAI 2025 · 11 citations
