Regist3R: Incremental Registration with Stereo Foundation Model
Sidun Liu, Wenyu Li, Peng Qiao, Yong Dou
摘要
Multi-view 3D reconstruction has remained an essential yet challenging problem in the field of computer vision. While DUSt3R and its successors have achieved breakthroughs in 3D reconstruction from unposed images, these methods exhibit significant limitations when scaling to multi-view scenarios, including high computational cost and cumulative error induced by global alignment. To address these challenges, we propose Regist3R, a novel stereo foundation model tailored for efficient and scalable incremental reconstruction. Regist3R leverages an incremental reconstruction paradigm, enabling large-scale 3D reconstructions from unordered and many-view image collections. We evaluate Regist3R on public datasets for camera pose estimation and 3D reconstruction. Our experiments demonstrate that Regist3R achieves comparable performance with optimization-based methods while significantly improving computational efficiency, and outperforms existing multi-view reconstruction models. Furthermore, to assess its performance in real-world applications, we introduce a challenging oblique aerial dataset which has long spatial spans and hundreds of views. The results highlight the effectiveness of Regist3R. We also demonstrate the first attempt to reconstruct large-scale scenes encompassing over thousands of views through pointmap-based foundation models, showcasing its potential for practical applications in large-scale 3D reconstruction tasks, including urban modeling, aerial mapping, and beyond.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper24
- Vision Transformers for Dense PredictionRené Ranftl, Alexey Bochkovskiy, Vladlen KoltunICCV 2021 · 被引用 2,647 次
- DROID-SLAM: Deep Visual SLAM for Monocular, Stereo, and RGB-D CamerasZachary Teed, Jia DengNeurIPS 2021 · 被引用 1,248 次
- Common Objects in 3D: Large-Scale Learning and Evaluation of Real-life 3D Category ReconstructionJeremy Reizenstein, Roman Shapovalov, Philipp Henzler, Luca Sbordone 等ICCV 2021 · 被引用 686 次
- ScanNet++: A High-Fidelity Dataset of 3D Indoor ScenesChandan Yeshwanth, Yueh-Cheng Liu, Matthias Nießner, Angela DaiICCV 2023 · 被引用 659 次
- Mega-NeRF: Scalable Construction of Large-Scale NeRFs for Virtual Fly- ThroughsHaithem Turki, Deva Ramanan, Mahadev SatyanarayananCVPR 2022 · 被引用 364 次
相关 Paper
- MUSt3R: Multi-view Network for Stereo 3D ReconstructionYohann Cabon, Lucas Stoffl, Leonid Antsfeld, Gabriela Csurka 等CVPR 2025
- Unposed Sparse Views Room Layout Reconstruction in the Age of Pretrain ModelYaxuan Huang, Xili Dai, Jianan Wang, Xianbiao Qi 等ICLR 2025
- DUSt3R: Geometric 3D Vision Made EasyShuzhe Wang, Vincent Leroy, Yohann Cabon, Boris Chidlovskii 等CVPR 2024 · 被引用 302 次
- MERG3R: A Divide-and-Conquer Approach to Large-Scale Neural Visual GeometryLeo Kaixuan Cheng, Abdus Shaikh, Ruofan Liang, Zhijie Wu 等CVPR 2026 · 被引用 3 次
- Mono3R: Exploiting Monocular Cues for Geometric 3D ReconstructionWenyu Li, Sidun Liu, Peng Qiao, Yong DouACM MM 2025 · 被引用 3 次
