Bidirectional Graph Reasoning Network for Panoptic Segmentation
Yangxin Wu, Gengwei Zhang, Yiming Gao, Xiajun Deng, Ke Gong, Xiaodan Liang, Liang Lin
摘要
Recent researches on panoptic segmentation resort to a single end-to-end network to combine the tasks of instance segmentation and semantic segmentation. However, prior models only unified the two related tasks at the architectural level via a multi-branch scheme or revealed the underlying correlation between them by unidirectional feature fusion, which disregards the explicit semantic and co-occurrence relations among objects and background. Inspired by the fact that context information is critical to recognize and localize the objects, and inclusive object details are significant to parse the background scene, we thus investigate on explicitly modeling the correlations between object and background to achieve a holistic understanding of an image in the panoptic segmentation task. We introduce a Bidirectional Graph Reasoning Network (BGRNet), which incorporates graph structure into the conventional panoptic segmentation network to mine the intra-modular and intermodular relations within and between foreground things and background stuff classes. In particular, BGRNet first constructs image-specific graphs in both instance and semantic segmentation branches that enable flexible reasoning at the proposal level and class level, respectively. To establish the correlations between separate branches and fully leverage the complementary relations between things and stuff, we propose a Bidirectional Graph Connection Module to diffuse information across branches in a learnable fashion. Experimental results demonstrate the superiority of our BGRNet that achieves the new state-of-the-art performance on challenging COCO and ADE20K panoptic segmentation benchmarks.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper24
- Per-Pixel Classification is Not All You Need for Semantic SegmentationBowen Cheng, Alexander G. Schwing, Alexander KirillovNeurIPS 2021 · 被引用 2,196 次
- Few-Shot Segmentation via Cycle-Consistent TransformerGengwei Zhang, Guoliang Kang, Yi Yang, Yunchao WeiNeurIPS 2021 · 被引用 282 次
- Panoptic SegFormer: Delving Deeper into Panoptic Segmentation with TransformersZhiqi Li, Wenhai Wang, Enze Xie, Zhiding Yu 等CVPR 2022 · 被引用 145 次
- Contrastive Language-Image Pre-Training with Knowledge GraphsXuran Pan, Tianzhu Ye, Dongchen Han, Shiji Song 等NeurIPS 2022 · 被引用 81 次
- Graph-based Spatial Transformer with Memory Replay for Multi-future Pedestrian Trajectory PredictionLihuan Li, Maurice Pagnucco, Yang SongCVPR 2022 · 被引用 76 次
相关 Paper
- BANet: Bidirectional Aggregation Network With Occlusion Handling for Panoptic SegmentationYifeng Chen, Guangchen Lin, Songyuan Li, Omar El Farouk Bourahla 等CVPR 2020
- Auto-Panoptic: Cooperative Multi-Component Architecture Search for Panoptic SegmentationYangxin Wu, Gengwei Zhang, Hang Xu, Xiaodan Liang 等NeurIPS 2020 · 被引用 21 次
- Panoptic, Instance and Semantic Relations: A Relational Context Encoder to Enhance Panoptic SegmentationShubhankar Borse, Hyojin Park, Hong Cai, Debasmit Das 等CVPR 2022 · 被引用 17 次
- SOGNet: Scene Overlap Graph Network for Panoptic SegmentationYibo Yang, Hongyang Li, Xia Li, Qijie Zhao 等AAAI 2020 · 被引用 64 次
- K-Net: Towards Unified Image SegmentationWenwei Zhang, Jiangmiao Pang, Kai Chen, Chen Change LoyNeurIPS 2021 · 被引用 500 次
