Learning Structure-From-Motion with Graph Attention Networks
Lucas Brynte, José Pedro Iglesias, Carl Olsson, Fredrik Kahl
摘要
In this paper we tackle the problem of learning Structurefrom-Motion (SfM) through the use of graph attention networks. SfM is a classic computer vision problem that is solved though iterative minimization of reprojection errors, referred to as Bundle Adjustment (BA), starting from a good initialization. In order to obtain a good enough initialization to BA, conventional methods rely on a sequence of sub-problems (such as pairwise pose estimation, pose averaging or triangulation) which provide an initial solution that can then be refined using BA. In this work we replace these sub-problems by learning a model that takes as input the 2D keypoints detected across multiple views, and outputs the corresponding camera poses and 3D keypoint coordinates. Our model takes advantage of graph neural networks to learn SfM-specific primitives, and we show that it can be used for fast inference of the reconstruction for new and unseen sequences. The experimental results show that the proposed model outperforms competing learning-based methods, and challenges COLMAP while having lower runtime.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- VGGSfM: Visual Geometry Grounded Deep Structure from MotionJianyuan Wang, Nikita Karaev, Christian Rupprecht, David NovotnýCVPR 2024 · 被引用 48 次
- Fast Encoder-Based 3D from Casual Videos via Point Track ProcessingYoni Kasten, Wuyue Lu, Haggai MaronNeurIPS 2024 · 被引用 16 次
- MuM: Multi-View Masked Image Modeling for 3D VisionDavid Nordström, Johan Edstedt, Fredrik Kahl, Georg BökmanCVPR 2026 · 被引用 6 次
- Global-Aware Edge Prioritization for Pose Graph InitializationTong Wei, Giorgos Tolias, Jiri Matas, Daniel BarathCVPR 2026 · 被引用 1 次
- Uncalibrated Structure from Motion on a SphereJonathan Ventura, Viktor Larsson, Fredrik KahlICCV 2025 · 被引用 1 次
它引用的顶会 Paper7
- How Attentive are Graph Attention Networks?Shaked Brody, Uri Alon, Eran YahavICLR 2022 · 被引用 1,717 次
- NerfingMVS: Guided Optimization of Neural Radiance Fields for Indoor Multi-view StereoYi Wei, Shaohui Liu, Yongming Rao, Wang Zhao 等ICCV 2021 · 被引用 286 次
- PoseDiffusion: Solving Pose Estimation via Diffusion-aided Bundle AdjustmentJianyuan Wang, Christian Rupprecht, David NovotnýICCV 2023 · 被引用 158 次
- Algebraic Characterization of Essential Matrices and Their Averaging in Multiview SettingsYoni Kasten, Amnon Geifman, Meirav Galun, Ronen BasriICCV 2019 · 被引用 35 次
- Deep Permutation Equivariant Structure from MotionDror Moran, Hodaya Koslowsky, Yoni Kasten, Haggai Maron 等ICCV 2021 · 被引用 21 次
相关 Paper
- Learning to Bundle-adjust: A Graph Network Approach to Faster Optimization of Bundle Adjustment for Vehicular SLAMTetsuya Tanaka, Yukihiro Sasagawa, Takayuki OkataniICCV 2021 · 被引用 9 次
- Light3R-SfM: Towards Feed-forward Structure-from-MotionSven Elflein, Qunjie Zhou, Laura Leal-TaixéCVPR 2025
- Pixel-Perfect Structure-from-Motion with Featuremetric RefinementPhilipp Lindenberger, Paul-Edouard Sarlin, Viktor Larsson, Marc PollefeysICCV 2021 · 被引用 266 次
- SuperGlue: Learning Feature Matching With Graph Neural NetworksPaul-Edouard Sarlin, Daniel DeTone, Tomasz Malisiewicz, Andrew RabinovichCVPR 2020
- Level-S2fM: Structure from Motion on Neural Level Set of Implicit SurfacesYuxi Xiao, Nan Xue, Tianfu Wu, Gui-Song XiaCVPR 2023
