StruMonoNet: Structure-Aware Monocular 3D Prediction
Zhenpei Yang, Li Erran Li, Qixing Huang
摘要
Monocular 3D prediction is one of the fundamental problems in 3D vision. Recent deep learning-based approaches have brought us exciting progress on this problem. However, existing approaches have predominantly focused on end-to-end depth and normal predictions, which do not fully utilize the underlying 3D environment’s geometric structures. This paper introduces StruMonoNet, which detects and enforces a planar structure to enhance pixel-wise predictions. StruMonoNet innovates in leveraging a hybrid representation that combines visual feature and a surfel representation for plane prediction. This formulation allows us to combine the power of visual feature learning and the flexibility of geometric representations in incorporating geometric relations. As a result, StruMonoNet can detect relations between planes such as adjacent planes, perpendicular planes, and parallel planes, all of which are beneficial for dense 3D prediction. Experimental results show that StruMonoNet considerably outperforms state-of-the-art approaches on NYUv2 and ScanNet.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- P3Depth: Monocular Depth Estimation with a Piecewise Planarity PriorVaishakh Patil, Christos Sakaridis, Alexander Liniger, Luc Van GoolCVPR 2022 · 被引用 144 次
- MVS2D: Efficient Multiview Stereo via Attention-Driven 2D ConvolutionsZhenpei Yang, Zhile Ren, Qi Shan, Qixing HuangCVPR 2022 · 被引用 43 次
它引用的顶会 Paper7
- Digging Into Self-Supervised Monocular Depth EstimationClément Godard, Oisin Mac Aodha, Michael Firman, Gabriel J. BrostowICCV 2019 · 被引用 2,416 次
- Enforcing Geometric Constraints of Virtual Normal for Depth PredictionWei Yin, Yifan Liu, Chunhua Shen, Youliang YanICCV 2019 · 被引用 487 次
- End-to-End Wireframe ParsingYichao Zhou, Haozhi Qi, Yi MaICCV 2019 · 被引用 190 次
- Learning to Reconstruct 3D Manhattan Wireframes From a Single ImageYichao Zhou, Haozhi Qi, Yuexiang Zhai, Qi Sun 等ICCV 2019 · 被引用 74 次
- FrameNet: Learning Local Canonical Frames of 3D Surfaces From a Single RGB ImageJingwei Huang, Yichao Zhou, Thomas A. Funkhouser, Leonidas J. GuibasICCV 2019 · 被引用 50 次
相关 Paper
- PlaneRAS: Learning Planar Primitives for 3D Plane RecoveryFang Zhang, Wenzhao Zheng, Linqing Zhao, Zelan Zhu 等ICCV 2025
- MonoDETR: Depth-guided Transformer for Monocular 3D Object DetectionRenrui Zhang, Han Qiu, Tai Wang, Ziyu Guo 等ICCV 2023 · 被引用 175 次
- Monocular 3D Object Detection with Decoupled Structured Polygon Estimation and Height-Guided Depth EstimationYingjie Cai, Buyu Li, Zeyu Jiao, Hongsheng Li 等AAAI 2020 · 被引用 100 次
- NDDepth: Normal-Distance Assisted Monocular Depth EstimationShuwei Shao, Zhongcai Pei, Weihai Chen, Xingming Wu 等ICCV 2023 · 被引用 76 次
- NeRD: Neural 3D Reflection Symmetry DetectorYichao Zhou, Shichen Liu, Yi MaCVPR 2021
