MV-RoMa: From Pairwise Matching into Multi-View Track Reconstruction
JongMin Lee, Seungyeop Kang, Sungjoo Yoo
Abstract
Establishing consistent correspondences across images is essential for 3D vision tasks such as structure-frommotion (SfM), yet most existing matchers operate in a pairwise manner, often producing fragmented and geometrically inconsistent tracks when their predictions are chained across views. We propose MV-RoMa, a multi-view dense matching model that jointly estimates dense correspondences from a source image to multiple co-visible targets.
Specifically, we design an efficient model architecture which avoids high computational cost of full cross-attention for multi-view feature interaction: (i) multi-view encoder that leverages pair-wise matching results as a geometric prior, and (ii) multi-view matching refiner that refines correspondences using pixel-wise attention. Additionally, we propose a post-processing strategy that integrates our model's consistent multi-view correspondences as highquality tracks for SfM. Across diverse and challenging benchmarks, MV-RoMa produces more reliable correspondences and substantially denser, more accurate 3D reconstructions than existing sparse and dense matching methods. Project page: https://icetea-cv.github. io/mv-roma/.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Builds on18
- 3D Gaussian Splatting for Real-Time Radiance Field RenderingBernhard Kerbl, Georgios Kopanas, Thomas Leimkühler, George DrettakisSIGGRAPH 2023 · 5,687 citations
- LightGlue: Local Feature Matching at Light SpeedPhilipp Lindenberger, Paul-Edouard Sarlin, Marc PollefeysICCV 2023 · 936 citations
- Depth Anything 3: Recovering the Visual Space from Any ViewsHaotong Lin, Sili Chen, Jun Hao Liew, Donny Y. Chen et al.ICLR 2026 · 720 citations
- DISK: Learning local features with policy gradientMichal J. Tyszkiewicz, Pascal Fua, Eduard TrullsNeurIPS 2020 · 652 citations
- Learning Two-View Correspondences and Geometry Using Order-Aware NetworkJiahui Zhang, Dawei Sun, Zixin Luo, Anbang Yao et al.ICCV 2019 · 362 citations
Related papers
- UFM: A Simple Path towards Unified Dense Correspondence with FlowYuchen Zhang, Nikhil Varma Keetha, Chenwei Lyu, Bhuvan Jhamb et al.NeurIPS 2025 · 40 citations
- Dense-SfM: Structure from Motion with Dense Consistent MatchingJongMin Lee, Sungjoo YooCVPR 2025
- RoMa: Robust Dense Feature MatchingJohan Edstedt, Qiyu Sun, Georg Bökman, Mårten Wadenbäck et al.CVPR 2024
- DOMR: Establishing Cross-View Segmentation via Dense Object MatchingJitong Liao, Yulu Gao, Shaofei Huang, Jialin Gao et al.ACM MM 2025
- DKM: Dense Kernelized Feature Matching for Geometry EstimationJohan Edstedt, Ioannis Athanasiadis, Mårten Wadenbäck, Michael FelsbergCVPR 2023
