PanelNet: Understanding 360 Indoor Environment via Panel Representation
Haozheng Yu, Lu He, Bing Jian, Weiwei Feng, Shan Liu
Abstract
Indoor 360 panoramas have two essential properties. (1) The panoramas are continuous and seamless in the horizontal direction. (2) Gravity plays an important role in indoor environment design. By leveraging these properties, we present PanelNet, a framework that understands indoor environments using a novel panel representation of 360 images. We represent an equirectangular projection (ERP) as consecutive vertical panels with corresponding 3D panel geometry. To reduce the negative impact of panoramic distortion, we incorporate a panel geometry embedding network that encodes both the local and global geometric features of a panel. To capture the geometric context in room design, we introduce Local2Global Transformer, which aggregates local information within a panel and panel-wise global context. It greatly improves the model performance with low training overhead. Our method outperforms existing methods on indoor 360 depth estimation and shows competitive results against stateof-the-art approaches on the task of indoor layout estimation and semantic segmentation.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers8
- Denoise and Align: Towards Source-Free UDA for Robust Panoramic Semantic SegmentationYaowen Chang, Zhen Cao, Xu Zheng, Xiaoxin Mi et al.CVPR 2026 · 4 citations
- PanoSplatt3R: Leveraging Perspective Pretraining for Generalized Unposed Wide-Baseline Panorama ReconstructionJiahui Ren, Mochu Xiang, Jiajun Zhu, Yuchao DaiICCV 2025 · 3 citations
- PanoContext-Former: Panoramic Total Scene Understanding with a TransformerYuan Dong, Chuan Fang, Liefeng Bo, Zilong Dong et al.CVPR 2024
- SO(3)-Equivariant ViT-Adapter for Data-Efficient Zero-Shot Sim-to-Real Indoor Panoramic Depth EstimationZiyan He, Qiudan Zhang, Lin Ma, Xu WangCVPR 2026
- PanDA: Towards Panoramic Depth Anything with Unlabeled Panoramas and Mobius Spatial AugmentationZidong Cao, Jinjing Zhu, Weiming Zhang, Hao Ai et al.CVPR 2025
Builds on14
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu et al.ICCV 2021 · 31,683 citations
- Vision Transformers for Dense PredictionRené Ranftl, Alexey Bochkovskiy, Vladlen KoltunICCV 2021 · 2,647 citations
- Tokens-to-Token ViT: Training Vision Transformers from Scratch on ImageNetLi Yuan, Yunpeng Chen, Tao Wang, Weihao Yu et al.ICCV 2021 · 2,462 citations
- Swin Transformer V2: Scaling Up Capacity and ResolutionZe Liu, Han Hu, Yutong Lin, Zhuliang Yao et al.CVPR 2022 · 2,138 citations
- Enforcing Geometric Constraints of Virtual Normal for Depth PredictionWei Yin, Yifan Liu, Chunhua Shen, Youliang YanICCV 2019 · 487 citations
Related papers
- SliceNet: Deep Dense Depth Estimation From a Single Indoor Panorama Using a Slice-Based RepresentationGiovanni Pintore, Marco Agus, Eva Almansa, Jens Schneider et al.CVPR 2021
- LGT-Net: Indoor Panoramic Room Layout Estimation with Geometry-Aware Transformer NetworkZhigang Jiang, Zhongzheng Xiang, Jinhua Xu, Ming ZhaoCVPR 2022 · 39 citations
- PanoSwin: a Pano-style Swin Transformer for Panorama UnderstandingZhixin Ling, Zhen Xing, Xiangdong Zhou, Manliang Cao et al.CVPR 2023
- LED2-Net: Monocular 360deg Layout Estimation via Differentiable Depth RenderingFu-En Wang, Yu-Hsuan Yeh, Min Sun, Wei-Chen Chiu et al.CVPR 2021
- PSMNet: Position-aware Stereo Merging Network for Room Layout EstimationHaiyan Wang, Will Hutchcroft, Yuguang Li, Zhiqiang Wan et al.CVPR 2022 · 20 citations
