Neural Reprojection Error: Merging Feature Learning and Camera Pose Estimation
Hugo Germain, Vincent Lepetit, Guillaume Bourmaud
摘要
Absolute camera pose estimation is usually addressed by sequentially solving two distinct subproblems: First a feature matching problem that seeks to establish putative 2D-3D correspondences, and then a Perspective-n-Point problem that minimizes, w.r.t. the camera pose, the sum of socalled Reprojection Errors (RE). We argue that generating putative 2D-3D correspondences 1) leads to an important loss of information that needs to be compensated as far as possible, within RE, through the choice of a robust loss and the tuning of its hyperparameters and 2) may lead to an RE that conveys erroneous data to the pose estimator. In this paper, we introduce the Neural Reprojection Error (NRE) as a substitute for RE. NRE allows to rethink the camera pose estimation problem by merging it with the feature learning problem, hence leveraging richer information than 2D-3D correspondences and eliminating the need for choosing a robust loss and its hyperparameters. Thus NRE can be used as training loss to learn image descriptors tailored for pose estimation. We also propose a coarse-to-fine optimization method able to very efficiently minimize a sum of NRE terms w.r.t. the camera pose. We experimentally demonstrate that NRE is a good substitute for RE as it significantly improves both the robustness and the accuracy of the camera pose estimate while being computationally and memory highly efficient. From a broader point of view, we believe this new way of merging deep learning and 3D geometry may be useful in other computer vision applications. Source code and model weights will be made available at hugogermain.com/nre.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper9
- Pixel-Perfect Structure-from-Motion with Featuremetric RefinementPhilipp Lindenberger, Paul-Edouard Sarlin, Viktor Larsson, Marc PollefeysICCV 2021 · 被引用 266 次
- Large-Scale Person Detection and Localization using Overhead Fisheye CamerasLu Yang, Liulei Li, Xueshi Xin, Yifan Sun 等ICCV 2023 · 被引用 35 次
- PUMP: Pyramidal and Uniqueness Matching Priors for Unsupervised Learning of Local DescriptorsJérôme Revaud, Vincent Leroy, Philippe Weinzaepfel, Boris ChidlovskiiCVPR 2022 · 被引用 16 次
- Visual Correspondence HallucinationHugo Germain, Vincent Lepetit, Guillaume BourmaudICLR 2022 · 被引用 11 次
- An Analytical Solution to Gauss-Newton Loss for Direct Image AlignmentSergei Solonets, Daniil Sinitsyn, Lukas von Stumberg, Nikita Araslanov 等ICLR 2024
它引用的顶会 Paper10
- DISK: Learning local features with policy gradientMichal J. Tyszkiewicz, Pascal Fua, Eduard TrullsNeurIPS 2020 · 被引用 652 次
- Learning Two-View Correspondences and Geometry Using Order-Aware NetworkJiahui Zhang, Dawei Sun, Zixin Luo, Anbang Yao 等ICCV 2019 · 被引用 362 次
- Neural-Guided RANSAC: Learning Where to Sample Model HypothesesEric Brachmann, Carsten RotherICCV 2019 · 被引用 282 次
- ELF: Embedded Localisation of Features in Pre-Trained CNNAssia Benbihi, Matthieu Geist, Cédric PradalierICCV 2019 · 被引用 30 次
- Reinforced Feature Points: Optimizing Feature Detection and Description for a High-Level TaskAritra Bhowmik, Stefan Gumhold, Carsten Rother, Eric BrachmannCVPR 2020
相关 Paper
- Neural Refinement for Absolute Pose Regression with Feature SynthesisShuai Chen, Yash Bhalgat, Xinghui Li, Jia-Wang Bian 等CVPR 2024
- CamNet: Coarse-to-Fine Retrieval for Camera Re-LocalizationMingyu Ding, Zhe Wang, Jiankai Sun, Jianping Shi 等ICCV 2019 · 被引用 163 次
- Focal Length and Object Pose Estimation via Render and CompareGeorgy Ponimatkin, Yann Labbé, Bryan C. Russell, Mathieu Aubry 等CVPR 2022 · 被引用 18 次
- NeMo: Neural Mesh Models of Contrastive Features for Robust 3D Pose EstimationAngtian Wang, Adam Kortylewski, Alan L. YuilleICLR 2021 · 被引用 53 次
- End-to-End Learnable Geometric Vision by Backpropagating PnP OptimizationBo Chen, Álvaro Parra, Jiewei Cao, Nan Li 等CVPR 2020
