Bidirectional Graph Reasoning Network for Panoptic Segmentation
Yangxin Wu, Gengwei Zhang, Yiming Gao, Xiajun Deng, Ke Gong, Xiaodan Liang, Liang Lin
Abstract
Recent researches on panoptic segmentation resort to a single end-to-end network to combine the tasks of instance segmentation and semantic segmentation. However, prior models only unified the two related tasks at the architectural level via a multi-branch scheme or revealed the underlying correlation between them by unidirectional feature fusion, which disregards the explicit semantic and co-occurrence relations among objects and background. Inspired by the fact that context information is critical to recognize and localize the objects, and inclusive object details are significant to parse the background scene, we thus investigate on explicitly modeling the correlations between object and background to achieve a holistic understanding of an image in the panoptic segmentation task. We introduce a Bidirectional Graph Reasoning Network (BGRNet), which incorporates graph structure into the conventional panoptic segmentation network to mine the intra-modular and intermodular relations within and between foreground things and background stuff classes. In particular, BGRNet first constructs image-specific graphs in both instance and semantic segmentation branches that enable flexible reasoning at the proposal level and class level, respectively. To establish the correlations between separate branches and fully leverage the complementary relations between things and stuff, we propose a Bidirectional Graph Connection Module to diffuse information across branches in a learnable fashion. Experimental results demonstrate the superiority of our BGRNet that achieves the new state-of-the-art performance on challenging COCO and ADE20K panoptic segmentation benchmarks.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext e646f7ff-d8e5-47b2-94ad-01c7820ad25fCited by top-tier papers24
- Per-Pixel Classification is Not All You Need for Semantic SegmentationBowen Cheng, Alexander G. Schwing, Alexander KirillovNeurIPS 2021 · 2,196 citations
- Few-Shot Segmentation via Cycle-Consistent TransformerGengwei Zhang, Guoliang Kang, Yi Yang, Yunchao WeiNeurIPS 2021 · 282 citations
- Panoptic SegFormer: Delving Deeper into Panoptic Segmentation with TransformersZhiqi Li, Wenhai Wang, Enze Xie, Zhiding Yu et al.CVPR 2022 · 145 citations
- Contrastive Language-Image Pre-Training with Knowledge GraphsXuran Pan, Tianzhu Ye, Dongchen Han, Shiji Song et al.NeurIPS 2022 · 81 citations
- Graph-based Spatial Transformer with Memory Replay for Multi-future Pedestrian Trajectory PredictionLihuan Li, Maurice Pagnucco, Yang SongCVPR 2022 · 76 citations
Related papers
- BANet: Bidirectional Aggregation Network With Occlusion Handling for Panoptic SegmentationYifeng Chen, Guangchen Lin, Songyuan Li, Omar El Farouk Bourahla et al.CVPR 2020
- Auto-Panoptic: Cooperative Multi-Component Architecture Search for Panoptic SegmentationYangxin Wu, Gengwei Zhang, Hang Xu, Xiaodan Liang et al.NeurIPS 2020 · 21 citations
- Panoptic, Instance and Semantic Relations: A Relational Context Encoder to Enhance Panoptic SegmentationShubhankar Borse, Hyojin Park, Hong Cai, Debasmit Das et al.CVPR 2022 · 17 citations
- SOGNet: Scene Overlap Graph Network for Panoptic SegmentationYibo Yang, Hongyang Li, Xia Li, Qijie Zhao et al.AAAI 2020 · 64 citations
- K-Net: Towards Unified Image SegmentationWenwei Zhang, Jiangmiao Pang, Kai Chen, Chen Change LoyNeurIPS 2021 · 500 citations
