Long-Tail Internet Photo Reconstruction
Yuan Li, Yuanbo Xiangli, Hadar Averbuch-Elor, Noah Snavely, Ruojin Cai
Abstract
Pretrained 𝜋 3 Ours Pretrained 𝜋 3 Ours Scene (sorted by #images) Registered images Total images #images per Scene Long Tail Duomo (Cagliari) -Crypt Calvaire de Plougonven Figure 1. Long-tail Internet photo reconstruction. Internet photo collections follow a long-tailed distribution. In the top plot, the x-axis represents scene index (sorted by image count) and the y-axis shows images per scene (scenes are drawn from MegaScenes [36], a dataset of Internet photo collections). The light blue curve plots the total number of Internet photos per scene, while the steel blue curve shows the size of the subset of photos that were successfully registered using SfM. The head of this distribution of photo collections represents well-photographed scenes; here, there are 6,985 scenes with >50 registered images. However, most photo collections are in the long tail of this distribution; here, 418,056 scenes with fewer than 50 registered photos. State-of-the-art methods often fail on scenes in this tail. In the lower half of the figure, we show two examples from the long tail, along with representative input images and the corresponding reconstructions.
On Calvaire de Plougonven, COLMAP doesn't register any image; on both Duomo (Cagliari)-Crypt and Calvaire de Plougonven, recent feed-forward reconstruction models like π 3 [44] produce poor results. We propose MegaDepth-X dataset and a strategy for mimicking long-tail camera distributions, on which fine-tuned models like π 3 exhibit better reconstruction robustness.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Builds on19
- LightGlue: Local Feature Matching at Light SpeedPhilipp Lindenberger, Paul-Edouard Sarlin, Marc PollefeysICCV 2023 · 936 citations
- Common Objects in 3D: Large-Scale Learning and Evaluation of Real-life 3D Category ReconstructionJeremy Reizenstein, Roman Shapovalov, Philipp Henzler, Luca Sbordone et al.ICCV 2021 · 686 citations
- DISK: Learning local features with policy gradientMichal J. Tyszkiewicz, Pascal Fua, Eduard TrullsNeurIPS 2020 · 652 citations
- DUSt3R: Geometric 3D Vision Made EasyShuzhe Wang, Vincent Leroy, Yohann Cabon, Boris Chidlovskii et al.CVPR 2024 · 302 citations
- Neural RGB-D Surface ReconstructionDejan Azinovic, Ricardo Martin-Brualla, Dan B. Goldman, Matthias Nießner et al.CVPR 2022 · 272 citations
Related papers
- Sparse-View Localization via Online Neural 3D RegressionLudvig Dillén, Magnus Oskarsson, Viktor LarssonCVPR 2026
- Extreme Rotation Estimation in the WildHana Bezalel, Dotan Ankri, Ruojin Cai, Hadar Averbuch-ElorCVPR 2025
- Geometry of Long-Tailed Representation Learning: Rebalancing Features for Skewed DistributionsLingjie Yi, Jiachen Yao, Weimin Lyu, Haibin Ling et al.ICLR 2025
- ULTRA-360: Unconstrained Dataset for Large-scale Temporal 3D Reconstruction across Altitudes and Omnidirectional ViewsXijun Liu, Zhaoliang Zhang, Yuxiang Guo, Yifan Zhou et al.ICLR 2026
- Towers of Babel: Combining Images, Language, and 3D Geometry for Learning Multimodal VisionXiaoshi Wu, Hadar Averbuch-Elor, Jin Sun, Noah SnavelyICCV 2021 · 26 citations
