Learning Dense Flow Field for Highly-accurate Cross-view Camera Localization
Zhenbo Song, Xianghui Ze, Jianfeng Lu, Yujiao Shi
Abstract
This paper addresses the problem of estimating the 3-DoF camera pose for a ground-level image with respect to a satellite image that encompasses the local surroundings. We propose a novel end-to-end approach that leverages the learning of dense pixel-wise flow fields in pairs of ground and satellite images to calculate the camera pose. Our approach differs from existing methods by constructing the feature metric at the pixel level, enabling full-image supervision for learning distinctive geometric configurations and visual appearances across views. Specifically, our method employs two distinct convolution networks for ground and satellite feature extraction. Then, we project the ground feature map to the bird's eye view (BEV) using a fixed camera height assumption to achieve preliminary geometric alignment. To further establish the content association between the BEV and satellite features, we introduce a residual convolution block to refine the projected BEV feature. Optical flow estimation is performed on the refined BEV feature map and the satellite feature map using flow decoder networks based on RAFT. After obtaining dense flow correspondences, we apply the least square method to filter matching inliers and regress the ground camera pose. Extensive experiments demonstrate significant improvements compared to state-of-the-art methods. Notably, our approach reduces the median localization error by 89%, 19%, 80%, and 35% on the KITTI, Ford multi-AV, VIGOR, and Oxford RobotCar datasets, respectively.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers10
- BevSplat: Resolving Height Ambiguity via Feature-Based Gaussian Primitives for Weakly-Supervised Cross-View LocalizationQiwei Wang, Shaoxun Wu, Yujiao ShiNeurIPS 2025 · 10 citations
- GeoDistill: Geometry-Guided Self-Distillation for Weakly Supervised Cross-View LocalizationShaowen Tong, Zimin Xia, Alexandre Alahi, Xuming He et al.ICCV 2025 · 3 citations
- Leveraging BEV Paradigm for Ground-to-Aerial Image SynthesisJunyan Ye, Jun He, Weijia Li, Zhutao Lv et al.ICCV 2025 · 2 citations
- Loc: Interpretable Cross-View Localization via Depth-Lifted Local Feature MatchingZimin Xia, Chenghao Xu, Alexandre AlahiICLR 2026 · 1 citation
- GeoFlow: Real-Time Fine-Grained Cross-View Geolocalization via Iterative Flow PredictionAyesh Abu Lehyeh, Xiaohan Zhang, Ahmad Arrabi, Waqas Sultani et al.CVPR 2026 · 1 citation
Builds on7
- Cross-view Geo-localization with Layer-to-Layer TransformerHongji Yang, Xiufan Lu, Yingying ZhuNeurIPS 2021 · 231 citations
- Beyond Cross-view Image Retrieval: Highly Accurate Vehicle Localization Using Satellite ImageYujiao Shi, Hongdong LiCVPR 2022 · 81 citations
- SliceMatch: Geometry-Guided Aggregation for Cross-View Pose EstimationTed de Vries Lentsch, Zimin Xia, Holger Caesar, Julian F. P. KooijCVPR 2023
- SuperGlue: Learning Feature Matching With Graph Neural NetworksPaul-Edouard Sarlin, Daniel DeTone, Tomasz Malisiewicz, Andrew RabinovichCVPR 2020
- VIGOR: Cross-View Image Geo-Localization Beyond One-to-One RetrievalSijie Zhu, Taojiannan Yang, Chen ChenCVPR 2021
Related papers
- FG^2: Fine-Grained Cross-View Localization by Fine-Grained Feature MatchingZimin Xia, Alexandre AlahiCVPR 2025
- VGA: Empowering Aerial-Ground Localization by Visual Geometry AlignmentTao Jun Lin, Yujiao Shi, Hongdong LiCVPR 2026
- CamLiFlow: Bidirectional Camera-LiDAR Fusion for Joint Optical Flow and Scene Flow EstimationHaisong Liu, Tao Lu, Yihui Xu, Jia Liu et al.CVPR 2022 · 64 citations
- VIRD: View-Invariant Representation through Dual-Axis Transformation for Cross-View Pose EstimationJuhye Park, Wooju Lee, Dasol Hong, Changki Sung et al.CVPR 2026
- EP2P-Loc: End-to-End 3D Point to 2D Pixel Localization for Large-Scale Visual LocalizationMinjung Kim, Junseo Koo, Gunhee KimICCV 2023 · 22 citations
