Self-Supervised Pretraining for RGB-D Salient Object Detection
Xiaoqi Zhao, Youwei Pang, Lihe Zhang, Huchuan Lu, Xiang Ruan
摘要
Existing CNNs-Based RGB-D salient object detection (SOD) networks are all required to be pretrained on the ImageNet to learn the hierarchy features which helps provide a good initialization. However, the collection and annotation of largescale datasets are time-consuming and expensive. In this paper, we utilize self-supervised representation learning (SSL) to design two pretext tasks: the cross-modal auto-encoder and the depth-contour estimation. Our pretext tasks require only a few and unlabeled RGB-D datasets to perform pretraining, which makes the network capture rich semantic contexts and reduce the gap between two modalities, thereby providing an effective initialization for the downstream task. In addition, for the inherent problem of cross-modal fusion in RGB-D SOD, we propose a consistency-difference aggregation (CDA) module that splits a single feature fusion into multi-path fusion to achieve an adequate perception of consistent and differential information. The CDA module is general and suitable for cross-modal and cross-level feature fusion. Extensive experiments on six benchmark datasets show that our self-supervised pretrained model performs favorably against most state-of-the-art methods pretrained on Im-ageNet. The source code will be publicly available at https: //github.com/Xiaoqi-Zhao-DLUT/SSLSOD .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper7
- Zoom In and Out: A Mixed-scale Triplet Network for Camouflaged Object DetectionYouwei Pang, Xiaoqi Zhao, Tian-Zhu Xiang, Lihe Zhang 等CVPR 2022 · 被引用 417 次
- Promoting Saliency From Depth: Deep Unsupervised RGB-D Saliency DetectionWei Ji, Jingjing Li, Qi Bi, Chuan Guo 等ICLR 2022 · 被引用 46 次
- Spider: A Unified Framework for Context-dependent Concept SegmentationXiaoqi Zhao, Youwei Pang, Wei Ji, Baicheng Sheng 等ICML 2024 · 被引用 21 次
- Multi-View Aggregation Network for Dichotomous Image SegmentationQian Yu, Xiaoqi Zhao, Youwei Pang, Lihe Zhang 等CVPR 2024 · 被引用 13 次
- Self-supervised Pre-training for Mirror DetectionJiaying Lin, Rynson W. H. LauICCV 2023 · 被引用 9 次
它引用的顶会 Paper13
- Depth-Induced Multi-Scale Recurrent Attention Network for Saliency DetectionYongri Piao, Wei Ji, Jingjing Li, Miao Zhang 等ICCV 2019 · 被引用 450 次
- Scaling and Benchmarking Self-Supervised Visual Representation LearningPriya Goyal, Dhruv Mahajan, Abhinav Gupta, Ishan MisraICCV 2019 · 被引用 429 次
- Depth Quality-Inspired Feature Manipulation for Efficient RGB-D Salient Object DetectionWenbo Zhang, Ge-Peng Ji, Zhuo Wang, Keren Fu 等ACM MM 2021 · 被引用 140 次
- Joint Semantic Mining for Weakly Supervised RGB-D Salient Object DetectionJingjing Li, Wei Ji, Qi Bi, Cheng Yan 等NeurIPS 2021 · 被引用 56 次
- You Only Infer Once: Cross-Modal Meta-Transfer for Referring Video Object SegmentationDezhuang Li, Ruoqi Li, Lijun Wang, Yifan Wang 等AAAI 2022 · 被引用 54 次
相关 Paper
- RGB-D Salient Object Detection via 3D Convolutional Neural NetworksQian Chen, Ze Liu, Yi Zhang, Keren Fu 等AAAI 2021 · 被引用 171 次
- Cross-modality Discrepant Interaction Network for RGB-D Salient Object DetectionChen Zhang, Runmin Cong, Qinwei Lin, Lin Ma 等ACM MM 2021 · 被引用 116 次
- CoMAE: Single Model Hybrid Pre-training on Small-Scale RGB-D DatasetsJiange Yang, Sheng Guo, Gangshan Wu, Limin WangAAAI 2023 · 被引用 12 次
- Specificity-preserving RGB-D Saliency DetectionTao Zhou, Huazhu Fu, Geng Chen, Yi Zhou 等ICCV 2021 · 被引用 210 次
- Self-Supervised Pretraining of 3D Features on any Point-CloudZaiwei Zhang, Rohit Girdhar, Armand Joulin, Ishan MisraICCV 2021 · 被引用 333 次
