FreeReg: Image-to-Point Cloud Registration Leveraging Pretrained Diffusion Models and Monocular Depth Estimators
Haiping Wang, Yuan Liu, Bing Wang, Yujing Sun, Zhen Dong, Wenping Wang, Bisheng Yang
摘要
Matching cross-modality features between images and point clouds is a fundamental problem for image-to-point cloud registration. However, due to the modality difference between images and points, it is difficult to learn robust and discriminative cross-modality features by existing metric learning methods for feature matching. Instead of applying metric learning on cross-modality data, we propose to unify the modality between images and point clouds by pretrained large-scale models first, and then establish robust correspondence within the same modality. We show that the intermediate features, called diffusion features, extracted by depth-to-image diffusion models are semantically consistent between images and point clouds, which enables the building of coarse but robust crossmodality correspondences. We further extract geometric features on depth maps produced by the monocular depth estimator. By matching such geometric features, we significantly improve the accuracy of the coarse correspondences produced by diffusion features. Extensive experiments demonstrate that without any training on the I2P registration task, direct utilization of both features produces accurate image-to-point cloud registration. On three public indoor and outdoor benchmarks, the proposed method averagely achieves a 20.6% improvement in Inlier Ratio, a 3.0× higher Inlier Number, and a 48.6% improvement in Registration Recall than existing state-of-the-arts. The code and additional results are available at https://whu-usi3dv.github.io/FreeReg/ .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper18
- Diff2I2P: Differentiable Image-to-Point Cloud Registration with Diffusion PriorJuncheng Mu, Chengwei Ren, Weixiang Zhang, Liang Pan 等ICCV 2025 · 被引用 9 次
- Vistadream: Sampling Multiview Consistent Images for Single-View Scene ReconstructionHaiping Wang, Yuan Liu, Ziwei Liu, Wenping Wang 等ICCV 2025 · 被引用 8 次
- MinCD-PnP: Learning 2D-3D Correspondences with Approximate Blind PnPPei An, Jiaqi Yang, Muyao Peng, You Yang 等ICCV 2025 · 被引用 5 次
- CA-I2P: Channel-Adaptive Registration Network with Global Optimal SelectionZhixin Cheng, Jiacheng Deng, Xinjun Li, Xiaotian Yin 等ICCV 2025 · 被引用 4 次
- RARE: Refine Any Registration of Pairwise Point Clouds via Zero-Shot LearningChengyu Zheng, Jin Huang, Honghua Chen, Mingqiang WeiICCV 2025 · 被引用 2 次
它引用的顶会 Paper41
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 被引用 13,211 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 被引用 11,743 次
相关 Paper
- Hg-I2P: Bridging Modalities for Generalizable Image-to-Point-Cloud Registration via Heterogeneous GraphsPei An, Junfeng Ding, Jiaqi Yang, Yulong Wang 等CVPR 2026 · 被引用 1 次
- Implicit Correspondence Learning for Image-to-Point Cloud RegistrationXinjun Li, Wenfei Yang, Jiacheng Deng, Zhixin Cheng 等CVPR 2025
- 2D3D-MATR: 2D-3D Matching Transformer for Detection-free Registration between Images and Point CloudsMinhao Li, Zheng Qin, Zhirui Gao, Renjiao Yi 等ICCV 2023 · 被引用 30 次
- DeepI2P: Image-to-Point Cloud Registration via Deep ClassificationJiaxin Li, Gim Hee LeeCVPR 2021
- Differentiable Registration of Images and LiDAR Point Clouds with VoxelPoint-to-Pixel MatchingJunsheng Zhou, Baorui Ma, Wenyuan Zhang, Yi Fang 等NeurIPS 2023 · 被引用 62 次
