Geometric Exploitation for Indoor Panoramic Semantic Segmentation
Dinh Duc Cao, Seok Joon Kim, Kyusung Cho
摘要
PAnoramic Semantic Segmentation (PASS) is an important task in computer vision, as it enables semantic understanding of a 360° environment. Currently, most of existing works have focused on addressing the distortion issues in 2D panoramic images without considering spatial properties of indoor scene. This restricts PASS methods in perceiving contextual attributes to deal with the ambiguity when working with monocular images. In this paper, we propose a novel approach for indoor panoramic semantic segmentation. Unlike previous works, we consider the panoramic image as a composition of segment groups: over-sampled segments , representing planar structures such as floors and ceilings, and under-sampled segments , representing other scene elements. To optimize each group, we first enhance over-sampled segments by jointly optimizing with a dense depth estimation task. Then, we introduce a transformer-based context module that aggregates different geometric representations of the scene, combined with a simple high-resolution branch, it serves as a robust hybrid decoder for estimating under-sampled segments , effectively preserving the resolution of predicted masks while leveraging various indoor geometric properties. Experimental results on both real-world (Stanford2D3DS, Matterport3D) and synthetic (Struc-tured3D) datasets demonstrate the robustness of our framework, by setting new state-of-the-arts in almost evaluations, The code and updated results are available at: https://github.com/caodinhduc/vertical_relative_distance .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper6
- Bending Reality: Distortion-aware Transformers for Adapting to Panoramic Semantic SegmentationJiaming Zhang, Kailun Yang, Chaoxiang Ma, Simon Reiß 等CVPR 2022 · 被引用 100 次
- ACDNet: Adaptively Combined Dilated Convolution for Monocular Panorama Depth EstimationChuanqing Zhuang, Zhengda Lu, Yiqun Wang, Jun Xiao 等AAAI 2022 · 被引用 73 次
- Look at the Neighbor: Distortion-aware Unsupervised Domain Adaptation for Panoramic Semantic SegmentationXu Zheng, Tianbo Pan, Yunhao Luo, Lin WangICCV 2023 · 被引用 46 次
- Capturing Omni-Range Context for Omnidirectional SegmentationKailun Yang, Jiaming Zhang, Simon Reiß, Xinxin Hu 等CVPR 2021
- Tangent Images for Mitigating Spherical DistortionMarc Eder, Mykhailo Shvets, John Lim, Jan-Michael FrahmCVPR 2020
相关 Paper
- PanoContext-Former: Panoramic Total Scene Understanding with a TransformerYuan Dong, Chuan Fang, Liefeng Bo, Zilong Dong 等CVPR 2024
- Denoise and Align: Towards Source-Free UDA for Robust Panoramic Semantic SegmentationYaowen Chang, Zhen Cao, Xu Zheng, Xiaoxin Mi 等CVPR 2026 · 被引用 4 次
- REL-SF4PASS: Panoramic Semantic Segmentation with REL Depth Representation and Spherical FusionXuewei Li, Xinghan Bao, Zhimin Chen, Xi LiCVPR 2026 · 被引用 3 次
- SDC-Depth: Semantic Divide-and-Conquer Network for Monocular Depth EstimationLijun Wang, Jianming Zhang, Oliver Wang, Zhe Lin 等CVPR 2020
- OmniSAM: Omnidirectional Segment Anything Model for UDA in Panoramic Semantic SegmentationDing Zhong, Xu Zheng, Chenfei Liao, Yuanhuiyi Lyu 等ICCV 2025 · 被引用 4 次
