Learning Multi-View Camera Relocalization With Graph Neural Networks
Fei Xue, Xin Wu, Shaojun Cai, Junqiu Wang
Abstract
We propose to construct a view graph to excavate the information of the whole given sequence for absolute camera pose estimation. Specifically, we harness GNNs to model the graph, allowing even non-consecutive frames to exchange information with each other. Rather than adopting the regular GNNs directly, we redefine the nodes, edges, and embedded functions to fit the relocalization task. Redesigned GNNs collaborate with CNNs in guiding knowledge propagation and feature extraction respectively to process multi-view high-dimensional image features iteratively at different levels. Besides, a general graph-based loss function beyond constraints between consecutive views is employed for training the network in an end-to-end fashion. Extensive experiments conducted on both indoor and outdoor datasets demonstrate that our method outperforms previous approaches especially in large-scale and challenging scenarios. Our code is publicly available ( https: //github.com/feixue94/grnet ).
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext d939a7d4-527b-4a51-878c-afbcfa1c02f5Cited by top-tier papers10
- PoGO-Net: Pose Graph Optimization with Graph Neural NetworksXinyi Li, Haibin LingICCV 2021 · 28 citations
- RobustLoc: Robust Camera Pose Regression in Challenging Driving EnvironmentsSijie Wang, Qiyu Kang, Rui She, Wee Peng Tay et al.AAAI 2023 · 27 citations
- Efficient Large-scale Localization by Global Instance RecognitionFei Xue, Ignas Budvytis, Daniel Olmeda Reino, Roberto CipollaCVPR 2022 · 19 citations
- VS-Net: Voting With Segmentation for Visual LocalizationZhaoyang Huang, Han Zhou, Yijin Li, Bangbang Yang et al.CVPR 2021
- Scene-agnostic Pose Regression for Visual LocalizationJunwei Zheng, Ruiping Liu, Yufan Chen, Zhenfang Chen et al.CVPR 2025
Builds on7
- Exploiting Spatial-Temporal Relationships for 3D Pose Estimation via Graph Convolutional NetworksYujun Cai, Liuhao Ge, Jun Liu, Jianfei Cai et al.ICCV 2019 · 504 citations
- Zero-Shot Video Object Segmentation via Attentive Graph Neural NetworksWenguan Wang, Xiankai Lu, Jianbing Shen, David J. Crandall et al.ICCV 2019 · 294 citations
- CamNet: Coarse-to-Fine Retrieval for Camera Re-LocalizationMingyu Ding, Zhe Wang, Jiankai Sun, Jianping Shi et al.ICCV 2019 · 163 citations
- SceneGraphNet: Neural Message Passing for 3D Indoor Scene AugmentationYang Zhou, Zachary While, Evangelos KalogerakisICCV 2019 · 109 citations
- Local Supports Global: Deep Camera Relocalization With Sequence EnhancementFei Xue, Xin Wang, Zike Yan, Qiuyuan Wang et al.ICCV 2019 · 57 citations
Related papers
- Keypoint Message Passing for Video-Based Person Re-identificationDi Chen, Andreas Doering, Shanshan Zhang, Jian Yang et al.AAAI 2022 · 25 citations
- Graph and Temporal Convolutional Networks for 3D Multi-person Pose Estimation in Monocular VideosYu Cheng, Bo Wang, Bo Yang, Robby T. TanAAAI 2021 · 55 citations
- View-GCN: View-Based Graph Convolutional Network for 3D Shape AnalysisXin Wei, Ruixuan Yu, Jian SunCVPR 2020
- DiffGlue: Diffusion-Aided Image Feature MatchingShihua Zhang, Jiayi MaACM MM 2024 · 4 citations
- Global-Aware Edge Prioritization for Pose Graph InitializationTong Wei, Giorgos Tolias, Jiri Matas, Daniel BarathCVPR 2026 · 1 citation
