Label Decoupling Framework for Salient Object Detection
Jun Wei, Shuhui Wang, Zhe Wu, Chi Su, Qingming Huang, Qi Tian
Abstract
To get more accurate saliency maps, recent methods mainly focus on aggregating multi-level features from fully convolutional network (FCN) and introducing edge information as auxiliary supervision. Though remarkable progress has been achieved, we observe that the closer the pixel is to the edge, the more difficult it is to be predicted, because edge pixels have a very imbalance distribution. To address this problem, we propose a label decoupling framework (LDF) which consists of a label decoupling (LD) procedure and a feature interaction network (FIN). LD explicitly decomposes the original saliency map into body map and detail map, where body map concentrates on center areas of objects and detail map focuses on regions around edges. Detail map works better because it involves much more pixels than traditional edge supervision. Different from saliency map, body map discards edge pixels and only pays attention to center areas. This successfully avoids the distraction from edge pixels during training. Therefore, we employ two branches in FIN to deal with body map and detail map respectively. Feature interaction (FI) is designed to fuse the two complementary branches to predict the saliency map, which is then used to refine the two branches again. This iterative refinement is helpful for learning better representations and more precise saliency maps. Comprehensive experiments on six benchmark datasets demonstrate that LDF outperforms state-of-the-art approaches on different evaluation metrics. Codes can be found at https: //github.com/weijun88/LDF .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 75ea22c9-cbdb-48c9-8b08-edacb4d0dee0Cited by top-tier papers38
- Visual Saliency TransformerNian Liu, Ni Zhang, Kaiyuan Wan, Ling Shao et al.ICCV 2021 · 473 citations
- I Can Find You! Boundary-Guided Separated Attention Network for Camouflaged Object DetectionHongwei Zhu, Peng Li, Haoran Xie, Xuefeng Yan et al.AAAI 2022 · 242 citations
- Locate Globally, Segment Locally: A Progressive Architecture With Knowledge Review Network for Salient Object DetectionBinwei Xu, Haoran Liang, Ronghua Liang, Peng ChenAAAI 2021 · 186 citations
- VisionLLM v2: An End-to-End Generalist Multimodal Large Language Model for Hundreds of Vision-Language TasksJiannan Wu, Muyan Zhong, Sen Xing, Zeqiang Lai et al.NeurIPS 2024 · 179 citations
- Learning Generative Vision Transformer with Energy-Based Latent Space for Saliency PredictionJing Zhang, Jianwen Xie, Nick Barnes, Ping LiNeurIPS 2021 · 117 citations
Builds on3
- EGNet: Edge Guidance Network for Salient Object DetectionJiaxing Zhao, Jiang-Jiang Liu, Deng-Ping Fan, Yang Cao et al.ICCV 2019 · 1,054 citations
- Stacked Cross Refinement Network for Edge-Aware Salient Object DetectionZhe Wu, Li Su, Qingming HuangICCV 2019 · 374 citations
- Selectivity or Invariance: Boundary-Aware Salient Object DetectionJinming Su, Jia Li, Yu Zhang, Changqun Xia et al.ICCV 2019 · 192 citations
Related papers
- Pyramidal Feature Shrinking for Salient Object DetectionMingcan Ma, Changqun Xia, Jia LiAAAI 2021 · 180 citations
- F³Net: Fusion, Feedback and Focus for Salient Object DetectionJun Wei, Shuhui Wang, Qingming HuangAAAI 2020 · 837 citations
- Global Context-Aware Progressive Aggregation Network for Salient Object DetectionZuyao Chen, Qianqian Xu, Runmin Cong, Qingming HuangAAAI 2020 · 481 citations
- Select, Supplement and Focus for RGB-D Saliency DetectionMiao Zhang, Weisong Ren, Yongri Piao, Zhengkun Rong et al.CVPR 2020
- Interactive Two-Stream Decoder for Accurate and Fast Saliency DetectionHuajun Zhou, Xiaohua Xie, Jian-Huang Lai, Zixuan Chen et al.CVPR 2020
