Bending Reality: Distortion-aware Transformers for Adapting to Panoramic Semantic Segmentation
Jiaming Zhang, Kailun Yang, Chaoxiang Ma, Simon Reiß, Kunyu Peng, Rainer Stiefelhagen
摘要
Panoramic images with their 360 • directional view encompass exhaustive information about the surrounding space, providing a rich foundation for scene understanding. To unfold this potential in the form of robust panoramic segmentation models, large quantities of expensive, pixelwise annotations are crucial for success. Such annotations are available, but predominantly for narrow-angle, pinholecamera images which, off the shelf, serve as sub-optimal resources for training panoramic models. Distortions and the distinct image-feature distribution in 360 • panoramas impede the transfer from the annotation-rich pinhole domain and therefore come with a big dent in performance. To get around this domain difference and bring together semantic annotations from pinhole-and 360 • surround-visuals, we propose to learn object deformations and panoramic image distortions in the Deformable Patch Embedding (DPE) and Deformable MLP (DMLP) components which blend into our Transformer for PAnoramic Semantic Segmentation (Trans4PASS) model. Finally, we tie together shared semantics in pinhole-and panoramic feature embeddings by generating multi-scale prototype features and aligning them in our Mutual Prototypical Adaptation (MPA) for unsupervised domain adaptation. On the indoor Stan-ford2D3D dataset, our Trans4PASS with MPA maintains comparable performance to fully-supervised state-of-thearts, cutting the need for over 1, 400 labeled panoramas. On the outdoor DensePASS dataset, we break state-of-theart by 14.39% mIoU and set the new bar at 56.38%. 1
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper21
- Look at the Neighbor: Distortion-aware Unsupervised Domain Adaptation for Panoramic Semantic SegmentationXu Zheng, Tianbo Pan, Yunhao Luo, Lin WangICCV 2023 · 被引用 46 次
- Sat2Density: Faithful Density Learning from Satellite-Ground Image PairsMing Qian, Jincheng Xiong, Gui-Song Xia, Nan XueICCV 2023 · 被引用 29 次
- SphereDiffusion: Spherical Geometry-Aware Distortion Resilient Diffusion ModelTao Wu, Xuewei Li, Zhongang Qi, Di Hu 等AAAI 2024 · 被引用 24 次
- Semantics, Distortion, and Style Matter: Towards Source-Free UDA for Panoramic SegmentationXu Zheng, Pengyuan Zhou, Athanasios V. Vasilakos, Lin WangCVPR 2024 · 被引用 16 次
- GoodSAM: Bridging Domain and Capacity Gaps via Segment Anything Model for Distortion-Aware Panoramic Semantic SegmentationWeiming Zhang, Yexin Liu, Xu Zheng, Lin WangCVPR 2024 · 被引用 14 次
它引用的顶会 Paper39
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu 等ICCV 2021 · 被引用 31,683 次
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- SegFormer: Simple and Efficient Design for Semantic Segmentation with TransformersEnze Xie, Wenhai Wang, Zhiding Yu, Anima Anandkumar 等NeurIPS 2021 · 被引用 9,661 次
- Training data-efficient image transformers & distillation through attentionHugo Touvron, Matthieu Cord, Matthijs Douze, Francisco Massa 等ICML 2021 · 被引用 8,974 次
- Deformable DETR: Deformable Transformers for End-to-End Object DetectionXizhou Zhu, Weijie Su, Lewei Lu, Bin Li 等ICLR 2021 · 被引用 7,353 次
相关 Paper
- Denoise and Align: Towards Source-Free UDA for Robust Panoramic Semantic SegmentationYaowen Chang, Zhen Cao, Xu Zheng, Xiaoxin Mi 等CVPR 2026 · 被引用 4 次
- Both Style and Distortion Matter: Dual-Path Unsupervised Domain Adaptation for Panoramic Semantic SegmentationXu Zheng, Jinjing Zhu, Yexin Liu, Zidong Cao 等CVPR 2023
- Geometric Exploitation for Indoor Panoramic Semantic SegmentationDinh Duc Cao, Seok Joon Kim, Kyusung ChoNeurIPS 2024 · 被引用 14 次
- Unlocking Constraints: Source-Free Occlusion-Aware Seamless SegmentationYihong Cao, Jiaming Zhang, Xu Zheng, Hao Shi 等ICCV 2025 · 被引用 4 次
- OmniSAM: Omnidirectional Segment Anything Model for UDA in Panoramic Semantic SegmentationDing Zhong, Xu Zheng, Chenfei Liao, Yuanhuiyi Lyu 等ICCV 2025 · 被引用 4 次
