Deep RGB-D Saliency Detection With Depth-Sensitive Attention and Automatic Multi-Modal Fusion
Peng Sun, Wenhu Zhang, Huanyu Wang, Songyuan Li, Xi Li
Abstract
RGB-D salient object detection (SOD) is usually formulated as a problem of classification or regression over two modalities, i.e., RGB and depth. Hence, effective RGB-D feature modeling and multi-modal feature fusion both play a vital role in RGB-D SOD. In this paper, we propose a depth-sensitive RGB feature modeling scheme using the depth-wise geometric prior of salient objects. In principle, the feature modeling scheme is carried out in a depth-sensitive attention module, which leads to the RGB feature enhancement as well as the background distraction reduction by capturing the depth geometry prior. Moreover, to perform effective multi-modal feature fusion, we further present an automatic architecture search approach for RGB-D SOD, which does well in finding out a feasible architecture from our specially designed multi-modal multi-scale search space. Extensive experiments on seven standard benchmarks demonstrate the effectiveness of the proposed approach against the state-of-the-art.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 15c82848-b415-4d0a-be03-b6295c4797fbCited by top-tier papers14
- TriTransNet: RGB-D Salient Object Detection with a Triplet Transformer Embedding NetworkZhengyi Liu, Yuan Wang, Zhengzheng Tu, Yun Xiao et al.ACM MM 2021 · 175 citations
- Learning Generative Vision Transformer with Energy-Based Latent Space for Saliency PredictionJing Zhang, Jianwen Xie, Nick Barnes, Ping LiNeurIPS 2021 · 117 citations
- DFormer: Rethinking RGBD Representation Learning for Semantic SegmentationBowen Yin, Xuying Zhang, Zhong-Yu Li, Li Liu et al.ICLR 2024 · 110 citations
- Self-Supervised Pretraining for RGB-D Salient Object DetectionXiaoqi Zhao, Youwei Pang, Lihe Zhang, Huchuan Lu et al.AAAI 2022 · 78 citations
- Point-aware Interaction and CNN-induced Refinement Network for RGB-D Salient Object DetectionRunmin Cong, Hongyu Liu, Chen Zhang, Wei Zhang et al.ACM MM 2023 · 71 citations
Builds on9
- Depth-Induced Multi-Scale Recurrent Attention Network for Saliency DetectionYongri Piao, Wei Ji, Jingjing Li, Miao Zhang et al.ICCV 2019 · 450 citations
- Auto-FPN: Automatic Network Architecture Adaptation for Object Detection Beyond ClassificationHang Xu, Lewei Yao, Zhenguo Li, Xiaodan Liang et al.ICCV 2019 · 197 citations
- Deep Multimodal Neural Architecture SearchZhou Yu, Yuhao Cui, Jun Yu, Meng Wang et al.ACM MM 2020 · 93 citations
- A2dele: Adaptive and Attentive Depth Distiller for Efficient RGB-D Salient Object DetectionYongri Piao, Zhengkun Rong, Miao Zhang, Weisong Ren et al.CVPR 2020
- Learning Selective Self-Mutual Attention for RGB-D Saliency DetectionNian Liu, Ni Zhang, Junwei HanCVPR 2020
Related papers
- RGB-D Salient Object Detection via 3D Convolutional Neural NetworksQian Chen, Ze Liu, Yi Zhang, Keren Fu et al.AAAI 2021 · 171 citations
- Is Depth Really Necessary for Salient Object Detection?Jiawei Zhao, Yifan Zhao, Jia Li, Xiaowu ChenACM MM 2020 · 71 citations
- MMNet: Multi-Stage and Multi-Scale Fusion Network for RGB-D Salient Object DetectionGuibiao Liao, Wei Gao, Qiuping Jiang, Ronggang Wang et al.ACM MM 2020 · 53 citations
- JL-DCF: Joint Learning and Densely-Cooperative Fusion Framework for RGB-D Salient Object DetectionKeren Fu, Deng-Ping Fan, Ge-Peng Ji, Qijun ZhaoCVPR 2020
- Feature Reintegration over Differential Treatment: A Top-down and Adaptive Fusion Network for RGB-D Salient Object DetectionMiao Zhang, Yu Zhang, Yongri Piao, Beiqi Hu et al.ACM MM 2020 · 51 citations
