ReTR: Modeling Rendering Via Transformer for Generalizable Neural Surface Reconstruction
Yixun Liang, Hao He, Yingcong Chen
Abstract
Generalizable neural surface reconstruction techniques have attracted great attention in recent years. However, they encounter limitations of low confidence depth distribution and inaccurate surface reasoning due to the oversimplified volume rendering process employed. In this paper, we present Reconstruction TRansformer (ReTR), a novel framework that leverages the transformer architecture to redesign the rendering process, enabling complex render interaction modeling. It introduces a learnable and utilizes the cross-attention mechanism to simulate the interaction of rendering process with sampled points and render the observed color. Meanwhile, by operating within a high-dimensional feature space rather than the color space, ReTR mitigates sensitivity to projected colors in source views. Such improvements result in accurate surface assessment with high confidence. We demonstrate the effectiveness of our approach on various datasets, showcasing how our method outperforms the current state-of-the-art approaches in terms of reconstruction quality and generalization ability. https://github.com/YixunLiang/ReTR.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext a854c78d-e312-4ac0-9d00-f998210b2a98Cited by top-tier papers12
- FatesGS: Fast and Accurate Sparse-View Surface Reconstruction Using Gaussian Splatting with Depth-Feature ConsistencyHan Huang, Yulun Wu, Chao Deng, Ge Gao et al.AAAI 2025 · 29 citations
- Sparfels: Fast Reconstruction from Sparse Unposed ImageryShubhendu Jena, Amine Ouasfi, Mae Younes, Adnane BoukhaymaICCV 2025 · 9 citations
- SparseRecon: Neural Implicit Surface Reconstruction from Sparse Views with Feature and Depth ConsistenciesLiang Han, Xu Zhang, Haichuan Song, Kanle Shi et al.ICCV 2025 · 5 citations
- RenderFormer: Transformer-based Neural Rendering of Triangle Meshes with Global IlluminationChong Zeng, Yue Dong, Pieter Peers, Hongzhi Wu et al.SIGGRAPH 2025 · 5 citations
- SurfelSplat: Learning Efficient and Generalizable Gaussian Surfel Representations for Sparse-View Surface ReconstructionChensheng Dai, Shengjun Zhang, Min Chen, Yueqi DuanNeurIPS 2025 · 3 citations
Builds on26
- NeuS: Learning Neural Implicit Surfaces by Volume Rendering for Multi-view ReconstructionPeng Wang, Lingjie Liu, Yuan Liu, Christian Theobalt et al.NeurIPS 2021 · 2,500 citations
- Volume Rendering of Neural Implicit SurfacesLior Yariv, Jiatao Gu, Yoni Kasten, Yaron LipmanNeurIPS 2021 · 1,421 citations
- MVSNeRF: Fast Generalizable Radiance Field Reconstruction from Multi-View StereoAnpei Chen, Zexiang Xu, Fuqiang Zhao, Xiaoshuai Zhang et al.ICCV 2021 · 1,024 citations
- Multiview Neural Surface Reconstruction by Disentangling Geometry and AppearanceLior Yariv, Yoni Kasten, Dror Moran, Meirav Galun et al.NeurIPS 2020 · 1,010 citations
- UNISURF: Unifying Neural Implicit Surfaces and Radiance Fields for Multi-View ReconstructionMichael Oechsle, Songyou Peng, Andreas GeigerICCV 2021 · 885 citations
Related papers
- Is Attention All That NeRF Needs?Mukund Varma T., Peihao Wang, Xuxi Chen, Tianlong Chen et al.ICLR 2023 · 6 citations
- Cross-View Geometric Collaboration for Generalizable Sparse View Neural Surface ReconstructionHang Yang, Le Hui, Jianjun Qian, Jian Yang et al.ACM MM 2025
- Multi-view 3D Reconstruction with TransformersDan Wang, Xinrui Cui, Xun Chen, Zhengxia Zou et al.ICCV 2021 · 111 citations
- GoLF-NRT: Integrating Global Context and Local Geometry for Few-Shot View SynthesisYou Wang, Li Fang, Hao Zhu, Fei Hu et al.CVPR 2025
- RnG: A Unified Transformer for Complete 3D Modeling from Partial ObservationsMochu Xiang, Zhelun Shen, Xuesong li, Jiahui Ren et al.CVPR 2026 · 2 citations
