A Global Depth-Range-Free Multi-View Stereo Transformer Network with Pose Embedding
Yitong Dong, Yijin Li, Zhaoyang Huang, Weikang Bian, Jingbo Liu, Hujun Bao, Zhaopeng Cui, Hongsheng Li, Guofeng Zhang
摘要
In this paper, we propose a novel multi-view stereo (MVS) framework that gets rid of the depth range prior. Unlike recent prior-free MVS methods that work in a pair-wise manner, our method simultaneously considers all the source images. Specifically, we introduce a Multi-view Disparity Attention (MDA) module to aggregate long-range context information within and across multi-view images. Considering the asymmetry of the epipolar disparity flow, the key to our method lies in accurately modeling multi-view geometric constraints. We integrate pose embedding to encapsulate information such as multi-view camera poses, providing implicit geometric constraints for multi-view disparity feature fusion dominated by attention. Additionally, we construct corresponding hidden states for each source image due to significant differences in the observation quality of the same pixel in the reference frame across multiple source frames. We explicitly estimate the quality of the current pixel corresponding to sampled points on the epipolar line of the source image and dynamically update hidden states through the uncertainty estimation module. Extensive results on the DTU dataset and Tanks&Temple benchmark demonstrate the effectiveness of our method. The code is available at our project page: https://zju3dv.github.io/GD-PoseMVS/.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- BlinkTrack: Feature Tracking Over 80 FPS via Events and ImagesYichen Shen, Yijin Li, Shuo Chen, Guanglin Li 等ICCV 2025 · 被引用 3 次
- One-Shot Refiner: Boosting Feed-forward Novel View Synthesis via One-Step DiffusionYitong Dong, Qi Zhang, Minchao Jiang, Zhiqiang Wu 等AAAI 2026 · 被引用 2 次
- Votesplat: Hough Voting Gaussian Splatting for 3D Scene UnderstandingMinchao Jiang, Shunyu Jia, Jiaming Gu, Xiaoyuan Lu 等ICCV 2025 · 被引用 1 次
- GeoCAD: Local Geometry-Controllable CAD Generation with Large Language ModelsZhanwei Zhang, Kaiyuan Liu, Junjie Liu, Wenxiao Wang 等NeurIPS 2025 · 被引用 1 次
它引用的顶会 Paper30
- P-MVSNet: Learning Patch-Wise Matching Confidence Aggregation for Multi-View StereoKeyang Luo, Tao Guan, Lili Ju, Haipeng Huang 等ICCV 2019 · 被引用 254 次
- TransMVSNet: Global Context-aware Multi-view Stereo Network with TransformersYikang Ding, Wentao Yuan, Qingtian Zhu, Haotian Zhang 等CVPR 2022 · 被引用 236 次
- AA-RMVSNet: Adaptive Aggregation Recurrent Multi-view Stereo NetworkZizhuang Wei, Qingtian Zhu, Chen Min, Yisong Chen 等ICCV 2021 · 被引用 193 次
- Rethinking Depth Estimation for Multi-View Stereo: A Unified RepresentationRui Peng, Rongjie Wang, Zhenyu Wang, Yawen Lai 等CVPR 2022 · 被引用 159 次
- Planar Prior Assisted PatchMatch Multi-View StereoQingshan Xu, Wenbing TaoAAAI 2020 · 被引用 154 次
相关 Paper
- Rethinking Disparity: A Depth Range Free Multi-View Stereo Based on DisparityQingsong Yan, Qiang Wang, Kaiyong Zhao, Bo Li 等AAAI 2023 · 被引用 21 次
- Attention-Aware Multi-View StereoKeyang Luo, Tao Guan, Lili Ju, Yuesong Wang 等CVPR 2020
- Point-Based Multi-View Stereo NetworkRui Chen, Songfang Han, Jing Xu, Hao SuICCV 2019 · 被引用 403 次
- Digging into Uncertainty in Self-supervised Multi-view StereoHongbin Xu, Zhipeng Zhou, Yali Wang, Wenxiong Kang 等ICCV 2021 · 被引用 68 次
- V-FUSE: Volumetric Depth Map Fusion with Long-Range ConstraintsNathaniel Burgdorfer, Philippos MordohaiICCV 2023 · 被引用 1 次
