DGC-GNN: Leveraging Geometry and Color Cues for Visual Descriptor-Free 2D-3D Matching
Shuzhe Wang, Juho Kannala, Daniel Barath
Abstract
Matching 2D keypoints in an image to a sparse 3D point cloud of the scene without requiring visual descriptors has garnered increased interest due to its low memory requirements, inherent privacy preservation, and reduced need for expensive 3D model maintenance compared to visual descriptor-based methods. However, existing algorithms of-ten compromise on performance, resulting in a significant de-terioration compared to their descriptor-based counterparts. In this paper, we introduce DGC-GNN, a novel algorithm that employs a global-to-local Graph Neural Network (GNN) that progressively exploits geometric and color cues to rep-resent keypoints, thereby improving matching accuracy. Our procedure encodes both Euclidean and angular relations at a coarse level, forming the geometric embedding to guide the point matching. We evaluate DGC-GNN on both indoor and outdoor datasets, demonstrating that it not only doubles the accuracy of the state-of-the-art visual descriptor-free algorithm but also substantially narrows the performance gap between descriptor-based and descriptor-free methods. <sup xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink">1</sup><sup xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink">1</sup>The code and trained models are available at: https://github.com/AaltoVision/DGC-GNN-release.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 1a88faa5-8f97-4e69-b92e-a507f5c99f14Cited by top-tier papers10
- LoD-Loc: Aerial Visual Localization using LoD 3D Map with Neural Wireframe AlignmentJuelin Zhu, Shen Yan, Long Wang, Shengyue Zhang et al.NeurIPS 2024 · 17 citations
- ULF-Loc: Unbiased Landmark Feature for Robust Visual Localization with 3D Gaussian SplattingYingdong Gu, Shaocheng Yan, Zhenjun Zhao, Yuan Kou et al.CVPR 2026 · 3 citations
- NormalLoc: Visual Localization on Textureless 3D Models using Surface NormalsJiro Abe, Gaku Nakano, Kazumine OguraICCV 2025 · 2 citations
- Revisiting Geometric Obfuscation with Dual Convergent Lines for Privacy-Preserving Image Queries in Visual LocalizationJeonggon Kim, Heejoon Moon, Je Hyeong HongCVPR 2026 · 1 citation
- ColabSfM: Collaborative Structure-from-Motion by Point Cloud RegistrationJohan Edstedt, André Mateus, Alberto JaenalCVPR 2025
Builds on13
- Transformers are RNNs: Fast Autoregressive Transformers with Linear AttentionAngelos Katharopoulos, Apoorv Vyas, Nikolaos Pappas, François FleuretICML 2020 · 2,665 citations
- Geometric Transformer for Fast and Robust Point Cloud RegistrationZheng Qin, Hao Yu, Changjian Wang, Yulan Guo et al.CVPR 2022 · 436 citations
- CoFiNet: Reliable Coarse-to-fine Correspondences for Robust PointCloud RegistrationHao Yu, Fu Li, Mahdi Saleh, Benjamin Busam et al.NeurIPS 2021 · 313 citations
- Pixel-Perfect Structure-from-Motion with Featuremetric RefinementPhilipp Lindenberger, Paul-Edouard Sarlin, Viktor Larsson, Marc PollefeysICCV 2021 · 266 citations
- Learning Multi-Scene Absolute Pose Regression with TransformersYoli Shavit, Ron Ferens, Yosi KellerICCV 2021 · 163 citations
Related papers
- SAG-GNN: Semantic-Aware Guided GNN for Descriptor-Free 2D-3D MatchingShihua Zhang, Tianhao Xu, Zizhuo Li, Qing Ma et al.CVPR 2026
- GlueStick: Robust Image Matching by Sticking Points and Lines TogetherRémi Pautrat, Iago Suárez, Yifan Yu, Marc Pollefeys et al.ICCV 2023 · 108 citations
- FC-GNN: Recovering Reliable and Accurate Correspondences from InterferencesHaobo Xu, Jun Zhou, Hua Yang, Renjie Pan et al.CVPR 2024 · 1 citation
- 2D3D-MATR: 2D-3D Matching Transformer for Detection-free Registration between Images and Point CloudsMinhao Li, Zheng Qin, Zhirui Gao, Renjiao Yi et al.ICCV 2023 · 30 citations
- LCD: Learned Cross-Domain Descriptors for 2D-3D MatchingQuang-Hieu Pham, Mikaela Angelina Uy, Binh-Son Hua, Duc Thanh Nguyen et al.AAAI 2020 · 94 citations
