Deep Object Co-Segmentation via Spatial-Semantic Network Modulation
Kaihua Zhang, Jin Chen, Bo Liu, Qingshan Liu
Abstract
Object co-segmentation is to segment the shared objects in multiple relevant images, which has numerous applications in computer vision. This paper presents a spatial and semantic modulated deep network framework for object cosegmentation. A backbone network is adopted to extract multi-resolution image features. With the multi-resolution features of the relevant images as input, we design a spatial modulator to learn a mask for each image. The spatial modulator captures the correlations of image feature descriptors via unsupervised learning. The learned mask can roughly localize the shared foreground object while suppressing the background. For the semantic modulator, we model it as a supervised image classification task. We propose a hierarchical second-order pooling module to transform the image features for classification use. The outputs of the two modulators manipulate the multi-resolution features by a shiftand-scale operation so that the features focus on segmenting co-object regions. The proposed model is trained end-to-end without any intricate post-processing. Extensive experiments on four image co-segmentation benchmark datasets demonstrate the superior accuracy of the proposed method compared to state-of-the-art methods. The codes are available at http://kaihuazhang.net/ .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers7
- Self-Supervised Object Localization with Joint Graph PartitionYukun Su, Guosheng Lin, Yun Hao, Yiwen Cao et al.AAAI 2022 · 17 citations
- Jack of All Tasks, Master of Many: Designing General-purpose Coarse-to-Fine Vision-Language ModelShraman Pramanick, Guangxing Han, Rui Hou, Sayan Nag et al.CVPR 2024 · 14 citations
- Contrastive Attention Maps for Self-supervised Co-localizationMinsong Ki, Youngjung Uh, Junsuk Choe, Hyeran ByunICCV 2021 · 11 citations
- Co-Salient Object Detection with Uncertainty-Aware Group Exchange-MaskingYang Wu, Huihui Song, Bo Liu, Kaihua Zhang et al.CVPR 2023
- CALICO: Part-Focused Semantic Co-Segmentation with Large Vision-Language ModelsKiet A. Nguyen, Adheesh Sunil Juvekar, Tianjiao Yu, Muntasir Wahed et al.CVPR 2025
Builds on1
Related papers
- Boat in the Sky: Background Decoupling and Object-aware Pooling for Weakly Supervised Semantic SegmentationJianjun Xu, Hongtao Xie, Hai Xu, Yuxin Wang et al.ACM MM 2022 · 13 citations
- Group-Wise Deep Object Co-Segmentation With Co-Attention Recurrent Neural NetworkBo Li, Zhengxing Sun, Qian Li, Yunjie Wu et al.ICCV 2019 · 62 citations
- Multi-scale Graph Fusion for Co-saliency DetectionRongyao Hu, Zhenyun Deng, Xiaofeng ZhuAAAI 2021 · 28 citations
- FOCUS: Towards Universal Foreground SegmentationZuyao You, Lingyu Kong, Lingchen Meng, Zuxuan WuAAAI 2025 · 9 citations
- GeoFree-CoSeg: Unsupervised Point Cloud-Image Cross-Modal Co-Segmentation Without Geometric AlignmentXin Duan, Xiabi Liu, Liyuan PanCVPR 2026
