ElasticMVS: Learning elastic part representation for self-supervised multi-view stereopsis
Jinzhi Zhang, Ruofan Tang, Zheng Cao, Jing Xiao, Ruqi Huang, Lu Fang
Abstract
Self-supervised multi-view stereopsis (MVS) attracts increasing attention for learning dense surface predictions from only a set of images without onerous ground-truth 3D training data for supervision. However, existing methods highly rely on the local photometric consistency, which fail to identify accurately dense correspondence in broad textureless or reflectant areas. In this paper, we show that geometric proximity such as surface connectedness and occlusion boundaries implicitly inferred from images could serve as reliable guidance for pixel-wise multi-view correspondences. With this insight, we present a novel elastic part representation, which encodes physically-connected part segmentations with elastically-varying scales, shapes and boundaries. Meanwhile, a self-supervised MVS framework namely ElasticMVS is proposed to learn the representation and estimate per-view depth following a part-aware propagation and evaluation scheme. Specifically, the pixel-wise part representation is trained by a contrastive learning-based strategy, which increases the representation compactness in geometrically concentrated areas and contrasts otherwise. ElasticMVS iteratively optimizes a part-level consistency loss and a surface smoothness loss, based on a set of depth hypotheses propagated from the geometrically concentrated parts. Extensive evaluations convey the superiority of ElasticMVS in the reconstruction completeness and accuracy, as well as the efficiency and scalability. Particularly, for the challenging large-scale reconstruction benchmark, ElasticMVS demonstrates significant performance gain over both the supervised and self-supervised approaches. Code is avaliable at https://thu-luvision.github.io.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 023ac18d-c1b3-49f0-b80a-9f0ac9e1594fCited by top-tier papers1
Ask how each one uses itBuilds on13
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 24,064 citations
- Unsupervised Learning of Visual Features by Contrasting Cluster AssignmentsMathilde Caron, Ishan Misra, Julien Mairal, Priya Goyal et al.NeurIPS 2020 · 5,249 citations
- Point-Based Multi-View Stereo NetworkRui Chen, Songfang Han, Jing Xu, Hao SuICCV 2019 · 403 citations
- TAPA-MVS: Textureless-Aware PAtchMatch Multi-View StereoAndrea Romanoni, Matteo MatteucciICCV 2019 · 95 citations
- Self-supervised Multi-view Stereo via Effective Co-Segmentation and Data-AugmentationHongbin Xu, Zhipeng Zhou, Yu Qiao, Wenxiong Kang et al.AAAI 2021 · 86 citations
Related papers
- Self-Supervised Multi-view Stereo via Adjacent Geometry Guided Volume CompletionLuoyuan Xu, Tao Guan, Yuesong Wang, Yawei Luo et al.ACM MM 2022 · 16 citations
- ParseMVS: Learning Primitive-aware Surface Representations for Sparse Multi-view StereopsisHaiyang Ying, Jinzhi Zhang, Yuzhe Chen, Zheng Cao et al.ACM MM 2022 · 4 citations
- MonoMVSNet: Monocular Priors Guided Multi-View Stereo NetworkJianfei Jiang, Qiankun Liu, Haochen Yu, Hongyuan Liu et al.ICCV 2025 · 3 citations
- Digging into Uncertainty in Self-supervised Multi-view StereoHongbin Xu, Zhipeng Zhou, Yali Wang, Wenxiong Kang et al.ICCV 2021 · 68 citations
- Just a Few Points are All You Need for Multi-view Stereo: A Novel Semi-supervised Learning Method for Multi-view StereoTaekyung Kim, Jaehoon Choi, Seokeon Choi, Dongki Jung et al.ICCV 2021 · 9 citations
