DenserNet: Weakly Supervised Visual Localization Using Multi-Scale Feature Aggregation
Dongfang Liu, Yiming Cui, Liqi Yan, Christos Mousas, Baijian Yang, Yingjie Victor Chen
摘要
In this work, we introduce a Denser Feature Network (DenserNet) for visual localization. Our work provides three principal contributions. First, we develop a convolutional neural network (CNN) architecture which aggregates feature maps at different semantic levels for image representations. Using denser feature maps, our method can produce more keypoint features and increase image retrieval accuracy. Second, our model is trained end-to-end without pixel-level annotation other than positive and negative GPS-tagged image pairs. We use a weakly supervised triplet ranking loss to learn discriminative features and encourage keypoint feature repeatability for image representation. Finally, our method is computationally efficient as our architecture has shared features and parameters during forwarding propagation. Our method is flexible and can be crafted on a light-weighted backbone architecture to achieve appealing efficiency with a small penalty on accuracy. Extensive experiment results indicate that our method sets a new state-of-the-art on four challenging large-scale localization benchmarks and three image retrieval benchmarks with the same level of supervision. The code is available at https://github.com/goodproj13/ DenserNet .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper12
- TF-Blender: Temporal Feature Blender for Video Object DetectionYiming Cui, Liqi Yan, Zhiwen Cao, Dongfang LiuICCV 2021 · 被引用 171 次
- EigenPlaces: Training Viewpoint Robust Models for Visual Place RecognitionGabriele Moreno Berton, Gabriele Trivigno, Barbara Caputo, Carlo MasoneICCV 2023 · 被引用 141 次
- Deep Visual Geo-localization BenchmarkGabriele Moreno Berton, Riccardo Mereu, Gabriele Trivigno, Carlo Masone 等CVPR 2022 · 被引用 80 次
- ClusterFomer: Clustering As A Universal Visual LearnerJames Liang, Yiming Cui, Qifan Wang, Tong Geng 等NeurIPS 2023 · 被引用 63 次
- Deep Homography Estimation for Visual Place RecognitionFeng Lu, Shuting Dong, Lijun Zhang, Bingxi Liu 等AAAI 2024 · 被引用 23 次
它引用的顶会 Paper1
相关 Paper
- Viewpoint Invariant Dense Matching for Visual GeolocalizationGabriele Moreno Berton, Carlo Masone, Valerio Paolicelli, Barbara CaputoICCV 2021 · 被引用 48 次
- EP2P-Loc: End-to-End 3D Point to 2D Pixel Localization for Large-Scale Visual LocalizationMinjung Kim, Junseo Koo, Gunhee KimICCV 2023 · 被引用 22 次
- MTLDesc: Looking Wider to Describe BetterChangwei Wang, Rongtao Xu, Yuyang Zhang, Shibiao Xu 等AAAI 2022 · 被引用 33 次
- Stochastic Attraction-Repulsion Embedding for Large Scale Image LocalizationLiu Liu, Hongdong Li, Yuchao DaiICCV 2019 · 被引用 123 次
- VS-Net: Voting With Segmentation for Visual LocalizationZhaoyang Huang, Han Zhou, Yijin Li, Bangbang Yang 等CVPR 2021
