Part-Aware Panoptic Segmentation
Daan de Geus, Panagiotis Meletis, Chenyang Lu, Xiaoxiao Wen, Gijs Dubbelman
Abstract
Part-aware panoptic segmentation (PPS) requires (a) that each foreground object and background region in an image is segmented and classified, and (b) that all parts within foreground objects are segmented, classified and linked to their parent object. Existing methods approach PPS by separately conducting object-level and part-level segmentation. However, their part-level predictions are not linked to individual parent objects. Therefore, their learning objective is not aligned with the PPS task objective, which harms the PPS performance. To solve this, and make more accurate PPS predictions, we propose Task-Aligned Part-aware Panoptic Segmentation (TAPPS). This method uses a set of shared queries to jointly predict (a) objectlevel segments, and (b) the part-level segments within those same objects. As a result, TAPPS learns to predict partlevel segments that are linked to individual parent objects, aligning the learning objective with the task objective, and allowing TAPPS to leverage joint object-part representations. With experiments, we show that TAPPS considerably outperforms methods that predict objects and parts separately, and achieves new state-of-the-art PPS results.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers16
- Hierarchical Open-vocabulary Universal Image SegmentationXudong Wang, Shufan Li, Konstantinos Kallidromitis, Yusuke Kato et al.NeurIPS 2023 · 74 citations
- Unifying Panoptic Segmentation for Autonomous DrivingOliver Zendel, Matthias Schörghuber, Bernhard Rainer, Markus Murschitz et al.CVPR 2022 · 49 citations
- Learning Hierarchical Image Segmentation For Recognition and By RecognitionTsung-Wei Ke, Sangwoo Mo, Stella X. YuICLR 2024 · 20 citations
- Auxiliary Learning as an Asymmetric Bargaining GameAviv Shamsian, Aviv Navon, Neta Glazer, Kenji Kawaguchi et al.ICML 2023 · 15 citations
- AIMS: All-Inclusive Multi-Level Segmentation for AnythingLu Qi, Jason Kuen, Weidong Guo, Jiuxiang Gu et al.NeurIPS 2023 · 9 citations
Builds on15
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu et al.ICCV 2021 · 31,683 citations
- SegFormer: Simple and Efficient Design for Semantic Segmentation with TransformersEnze Xie, Wenhai Wang, Zhiding Yu, Anima Anandkumar et al.NeurIPS 2021 · 9,661 citations
- A ConvNet for the 2020sZhuang Liu, Hanzi Mao, Chao-Yuan Wu, Christoph Feichtenhofer et al.CVPR 2022 · 6,782 citations
- Per-Pixel Classification is Not All You Need for Semantic SegmentationBowen Cheng, Alexander G. Schwing, Alexander KirillovNeurIPS 2021 · 2,196 citations
- Deep Hierarchical Semantic SegmentationLiulei Li, Tianfei Zhou, Wenguan Wang, Jianwu Li et al.CVPR 2022 · 181 citations
Related papers
- Task-Aligned Part-Aware Panoptic Segmentation Through Joint Object-Part RepresentationsDaan de Geus, Gijs DubbelmanCVPR 2024
- Towards Deeply Unified Depth-aware Panoptic Segmentation with Bi-directional Guidance LearningJunwen He, Yifan Wang, Lijun Wang, Huchuan Lu et al.ICCV 2023 · 11 citations
- LPSNet: A Lightweight Solution for Fast Panoptic SegmentationWeixiang Hong, Qingpei Guo, Wei Zhang, Jingdong Chen et al.CVPR 2021
- Slot-VPS: Object-centric Representation Learning for Video Panoptic SegmentationYi Zhou, Hui Zhang, Hana Lee, Shuyang Sun et al.CVPR 2022 · 20 citations
- BANet: Bidirectional Aggregation Network With Occlusion Handling for Panoptic SegmentationYifeng Chen, Guangchen Lin, Songyuan Li, Omar El Farouk Bourahla et al.CVPR 2020
