Adversarial Exploitation of Data Diversity Improves Visual Localization
Sihang Li, Siqi Tan, Bowen Chang, Jing Zhang, Chen Feng, Yiming Li
摘要
Visual localization, which estimates a camera's pose within a known scene, is a fundamental capability for autonomous systems. While absolute pose regression (APR) methods have shown promise for efficient inference, they often struggle with generalization. Recent approaches attempt to address this through data augmentation with varied viewpoints, yet they overlook a critical factor: appearance diversity. In this work, we identify appearance variation as the key to robust localization. Specifically, we first lift real 2D images into 3D Gaussian Splats with varying appearance and deblurring ability, enabling the synthesis of diverse training data that varies not just in poses but also in environmental conditions such as lighting and weather. To fully unleash the potential of the appearance-diverse data, we build a two-branch joint training pipeline with an adversarial discriminator to bridge the syn-to-real gap. Extensive experiments demonstrate that our approach significantly outperforms state-of-the-art methods, reducing translation and rotation errors by 50% and 33% on indoor datasets, and 38% and 44% on outdoor datasets. Most notably, our method shows remarkable robustness in dynamic driving scenarios under varying weather conditions and in day-to-night scenarios, where previous APR methods fail.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- Semantic-Guided Camera Ray Regression for Visual LocalizationYesheng Zhang, Xu ZhaoICCV 2025 · 被引用 2 次
- Rethinking Pose Refinement in 3D Gaussian Splatting under Pose Prior and Geometric UncertaintyMangyu Kong, Jaewon Lee, Seongwon Lee, Euntai KimCVPR 2026 · 被引用 1 次
- Simple but Effective Triplet-Based Compression Strategies for Compact Visual LocalizationTorsten Sattler, Zuzana KukelovaCVPR 2026
- Hierarchical Visual Relocalization with Nearest View Synthesis from Feature Gaussian SplattingHuaqi Tao, Bingxi Liu, Guangcheng Chen, Fulin Tang 等CVPR 2026
它引用的顶会 Paper20
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- 3D Gaussian Splatting for Real-Time Radiance Field RenderingBernhard Kerbl, Georgios Kopanas, Thomas Leimkühler, George DrettakisSIGGRAPH 2023 · 被引用 5,687 次
- VICReg: Variance-Invariance-Covariance Regularization for Self-Supervised LearningAdrien Bardes, Jean Ponce, Yann LeCunICLR 2022 · 被引用 1,226 次
- LightGlue: Local Feature Matching at Light SpeedPhilipp Lindenberger, Paul-Edouard Sarlin, Marc PollefeysICCV 2023 · 被引用 936 次
- DUSt3R: Geometric 3D Vision Made EasyShuzhe Wang, Vincent Leroy, Yohann Cabon, Boris Chidlovskii 等CVPR 2024 · 被引用 302 次
相关 Paper
- RobustLoc: Robust Camera Pose Regression in Challenging Driving EnvironmentsSijie Wang, Qiyu Kang, Rui She, Wee Peng Tay 等AAAI 2023 · 被引用 27 次
- AERGS-SLAM: Auto-Exposure-Robust Stereo 3D Gaussian Splatting SLAMZhiyu Zhou, Feng Hui, Yu LiuCVPR 2026
- GS-CPR: Efficient Camera Pose Refinement via 3D Gaussian SplattingChangkun Liu, Shuai Chen, Yash Bhalgat, Siyan Hu 等ICLR 2025 · 被引用 1 次
- DG-SLAM: Robust Dynamic Gaussian Splatting SLAM with Hybrid Pose OptimizationYueming Xu, Haochen Jiang, Zhongyang Xiao, Jianfeng Feng 等NeurIPS 2024 · 被引用 65 次
- GPVK-VL: Geometry-Preserving Virtual Keyframes for Visual Localization under Large Viewpoint ChangesYunxuan Li, Lei Fan, Xiaoying Xing, Jianxiong Zhou 等CVPR 2025
