Prior Guided Dropout for Robust Visual Localization in Dynamic Environments
Zhaoyang Huang, Yan Xu, Jianping Shi, Xiaowei Zhou, Hujun Bao, Guofeng Zhang
Abstract
Camera localization from monocular images has been a long-standing problem, but its robustness in dynamic environments is still not adequately addressed. Compared with classic geometric approaches, modern CNN-based methods (e.g. PoseNet) have manifested the reliability against illumination or viewpoint variations, but they still have the following limitations. First, foreground moving objects are not explicitly handled, which results in poor performance and instability in dynamic environments. Second, the output for each image is a point estimate without uncertainty quantification. In this paper, we propose a framework which can be generally applied to existing CNN-based pose regressors to improve their robustness in dynamic environments. The key idea is a prior guided dropout module coupled with a self-attention module which can guide CNNs to ignore foreground objects during both training and inference. Additionally, the dropout module enables the pose regressor to output multiple hypotheses from which the uncertainty of pose estimates can be quantified and leveraged in the following uncertainty-aware pose graph optimization to improve the robustness further. We achieve an average accuracy of 9.98m/3.63 • on RobotCar dataset, which outperforms the state-of-the-art method by 62.97%/47.08%. The source code of our implementation is available at https://github.com/zju3dv/RVL-Dynamic .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers11
- AtLoc: Attention Guided Camera LocalizationBing Wang, Changhao Chen, Chris Xiaoxuan Lu, Peijun Zhao et al.AAAI 2020 · 189 citations
- RobustLoc: Robust Camera Pose Regression in Challenging Driving EnvironmentsSijie Wang, Qiyu Kang, Rui She, Wee Peng Tay et al.AAAI 2023 · 27 citations
- LiSA: LiDAR Localization with Semantic AwarenessBochun Yang, Zijun Li, Wen Li, Zhipeng Cai et al.CVPR 2024 · 9 citations
- NopeRoomGS: Indoor 3D Gaussian Splatting Optimization without Camera Pose InputWenbo Li, Yan Xu, Mingde Yao, Fengjie Liang et al.NeurIPS 2025 · 1 citation
- LEADER: Learning Reliable Local-to-Global Correspondences for LiDAR RelocalizationJianshi Wu, Minghang Zhu, dq Liu, Wen Li et al.CVPR 2026 · 1 citation
Related papers
- Learning Camera Localization via Dense Scene MatchingShitao Tang, Chengzhou Tang, Rui Huang, Siyu Zhu et al.CVPR 2021
- MonoRUn: Monocular 3D Object Detection by Reconstruction and Uncertainty PropagationHansheng Chen, Yuyao Huang, Wei Tian, Zhong Gao et al.CVPR 2021
- PoGO-Net: Pose Graph Optimization with Graph Neural NetworksXinyi Li, Haibin LingICCV 2021 · 28 citations
- Digging into Uncertainty in Self-supervised Multi-view StereoHongbin Xu, Zhipeng Zhou, Yali Wang, Wenxiong Kang et al.ICCV 2021 · 68 citations
- GCCN: Geometric Constraint Co-attention Network for 6D Object Pose EstimationYongming Wen, Yiquan Fang, Junhao Cai, Kimwa Tung et al.ACM MM 2021 · 13 citations
