NeRF-IBVS: Visual Servo Based on NeRF for Visual Localization and Navigation
Yuanze Wang, Yichao Yan, Dianxi Shi, Wenhan Zhu, Jianqiang Xia, Jeff Tan, Songchang Jin, Ke Gao, Xiaobo Li, Xiaokang Yang
Abstract
Visual localization is a fundamental task in computer vision and robotics. Training existing visual localization methods requires a large number of posed images to generalize to novel views, while state-of-the-art methods generally require ground truth 3D labels for supervision. However, acquiring a large number of posed images and 3D labels in the real world is challenging and costly. In this paper, we present a novel visual localization method that achieves accurate localization while using only a few posed images compared to other localization methods. To achieve this, we first use a few posed images with coarse pseudo-3D labels provided by NeRF to train a coordinate regression network. Then a coarse pose is estimated from the regression network with PNP. Finally, we use the image-based visual servo (IBVS) with the scene prior provided by NeRF for pose optimization. Furthermore, our method can provide effective navigation prior, which enables navigation based on IBVS without using custom markers and the depth sensor. Extensive experiments on 7-Scenes and 12-Scenes datasets demonstrate that our method outperforms state-of-the-art methods under the same setting, with only 5% to 25% training data. Furthermore, our framework can be naturally extended to the visual navigation task based on IBVS, and its effectiveness is verified in simulation experiments.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 98a15453-a044-4cbd-9409-953963bfd48cCited by top-tier papers6
- F-3DGS: Factorized Coordinates and Representations for 3D Gaussian SplattingXiangyu Sun, Joo Chan Lee, Daniel Rho, Jong Hwan Ko et al.ACM MM 2024 · 14 citations
- 3D Gaussian Splatting based Scene-independent Relocalization with Unidirectional and Bidirectional Feature FusionJunyi Wang, Yuze Wang, Wantong Duan, Meng Wang et al.NeurIPS 2025 · 1 citation
- An Effective Levelling Paradigm for Unlabeled ScenariosFangming Cui, Zhou Yu, Di Yang, Yuqiang Ren et al.NeurIPS 2025
- PlanaReLoc: Camera Relocalization in 3D Planar Primitives via Region-Based Structure MatchingHanqiao Ye, Yuzhou Liu, Yangdong Liu, Shuhan ShenCVPR 2026
- Enhancing Target-unspecific Tasks through a Features MatrixFangming Cui, Yonggang Zhang, Xuan Wang, Xinmei Tian et al.ICML 2025
Builds on12
- BARF: Bundle-Adjusting Neural Radiance FieldsChen-Hsuan Lin, Wei-Chiu Ma, Antonio Torralba, Simon LuceyICCV 2021 · 867 citations
- iMAP: Implicit Mapping and Positioning in Real-TimeEdgar Sucar, Shikun Liu, Joseph Ortiz, Andrew J. DavisonICCV 2021 · 834 citations
- NICE-SLAM: Neural Implicit Scalable Encoding for SLAMZihan Zhu, Songyou Peng, Viktor Larsson, Weiwei Xu et al.CVPR 2022 · 720 citations
- Nerfstudio: A Modular Framework for Neural Radiance Field DevelopmentMatthew Tancik, Ethan Weber, Evonne Ng, Ruilong Li et al.SIGGRAPH 2023 · 592 citations
- AtLoc: Attention Guided Camera LocalizationBing Wang, Changhao Chen, Chris Xiaoxuan Lu, Peijun Zhao et al.AAAI 2020 · 189 citations
Related papers
- PNeRFLoc: Visual Localization with Point-Based Neural Radiance FieldsBoming Zhao, Luwei Yang, Mao Mao, Hujun Bao et al.AAAI 2024 · 31 citations
- Pose-Free Neural Radiance Fields via Implicit Pose RegularizationJiahui Zhang, Fangneng Zhan, Yingchen Yu, Kunhao Liu et al.ICCV 2023 · 17 citations
- Reloc3r: Large-Scale Training of Relative Camera Pose Regression for Generalizable, Fast, and Accurate Visual LocalizationSiyan Dong, Shuzhe Wang, Shaohui Liu, Lulu Cai et al.CVPR 2025
- CROSSFIRE: Camera Relocalization On Self-Supervised Features from an Implicit RepresentationArthur Moreau, Nathan Piasco, Moussâb Bennehar, Dzmitry Tsishkou et al.ICCV 2023 · 62 citations
- CrossLoc: Scalable Aerial Localization Assisted by Multimodal Synthetic DataQi Yan, Jianhao Zheng, Simon Reding, Shanci Li et al.CVPR 2022 · 23 citations
