PlaneRecTR: Unified Query Learning for 3D Plane Recovery from a Single View
Jingjia Shi, Shuaifeng Zhi, Kai Xu
摘要
3D plane recovery from a single image can usually be divided into several subtasks of plane detection, segmentation, parameter estimation and possibly depth estimation. Previous works tend to solve it by either extending the RCNN-based segmentation network or the dense pixel embedding-based clustering framework. However, none of them tried to integrate above related subtasks into a unified framework but treated them separately and sequentially, which we suspect is potentially a main source of performance limitation for existing approaches. Motivated by this finding and the success of query-based learning in enriching reasoning among semantic entities, in this paper, we propose PlaneRecTR, a Transformer-based architecture, which for the first time unifies all subtasks related to single-view plane recovery with a single compact model. Extensive quantitative and qualitative experiments demonstrate that our proposed unified learning achieves mutual benefits across subtasks, obtaining a new state-ofthe-art performance on public ScanNet and NYUv2-Plane datasets. Codes are available at https://github. com/SJingjia/PlaneRecTR.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper11
- AlphaTablets: A Generic Plane Representation for 3D Planar Reconstruction from Monocular VideosYuze He, Wang Zhao, Shaohui Liu, Yubin Hu 等NeurIPS 2024 · 被引用 7 次
- PLANA3R: Zero-shot Metric Planar 3D Reconstruction via Feed-forward Planar SplattingChangkun Liu, Bin Tan, Zeran Ke, Shangzhan Zhang 等NeurIPS 2025 · 被引用 6 次
- Novel View Synthesis Under Large-Deviation Viewpoint for Autonomous DrivingXin Ma, Jiguang Zhang, Peng Lu, Shibiao Xu 等AAAI 2025 · 被引用 5 次
- Diorama: Unleashing Zero-Shot Single-View 3D Indoor Scene ModelingQirui Wu, Denys Iliash, Daniel Ritchie, Manolis Savva 等ICCV 2025 · 被引用 4 次
- AirPlanes: Accurate Plane Estimation via 3D-Consistent EmbeddingsJamie Watson, Filippo Aleotti, Mohamed Sayed, Zawar Qureshi 等CVPR 2024 · 被引用 3 次
它引用的顶会 Paper9
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu 等ICCV 2021 · 被引用 31,683 次
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- Training data-efficient image transformers & distillation through attentionHugo Touvron, Matthieu Cord, Matthijs Douze, Francisco Massa 等ICML 2021 · 被引用 8,974 次
- Per-Pixel Classification is Not All You Need for Semantic SegmentationBowen Cheng, Alexander G. Schwing, Alexander KirillovNeurIPS 2021 · 被引用 2,196 次
- In-Place Scene Labelling and Understanding with Implicit Scene RepresentationShuaifeng Zhi, Tristan Laidlow, Stefan Leutenegger, Andrew J. DavisonICCV 2021 · 被引用 551 次
相关 Paper
- PlaneTR: Structure-Guided Transformers for 3D Plane RecoveryBin Tan, Nan Xue, Song Bai, Tianfu Wu 等ICCV 2021 · 被引用 51 次
- PlaneRAS: Learning Planar Primitives for 3D Plane RecoveryFang Zhang, Wenzhao Zheng, Linqing Zhao, Zelan Zhu 等ICCV 2025
- Uni-3D: A Universal Model for Panoptic 3D Scene ReconstructionXiang Zhang, Zeyuan Chen, Fangyin Wei, Zhuowen TuICCV 2023 · 被引用 24 次
- Towards In-the-wild 3D Plane Reconstruction from a Single ImageJiachen Liu, Rui Yu, Sili Chen, Sharon X. Huang 等CVPR 2025
- PlanarRecon: Realtime 3D Plane Detection and Reconstruction from Posed Monocular VideosYiming Xie, Matheus Gadelha, Fengting Yang, Xiaowei Zhou 等CVPR 2022 · 被引用 32 次
