FSCDiff: Frequency-Spatial Entangled Conditional Diffusion model for Underwater Salient Object Detection
Hua Li, Gaowei Lin, Zhiyuan Li, Sam Kwong, Runmin Cong
Abstract
Salient object detection (SOD) plays a crucial role in image understanding and visual guidance. However, due to the complexity of underwater environments, the accuracy of underwater salient object detection is often low. To improve the accuracy and robustness of underwater salient object detection, different from the existing spatial domain aware RGB-D methods that rely on pixel-level probabilities, we propose a novel Fourier-Spatial Entangled Conditional Diffusion model (FSCDiff) for underwater salient object detection. The FSCDiff aims to address the insufficient representation and boundary shift issues in underwater salient object detection by leveraging Fourier-domain information and the powerful multi-step iterative generation capability of diffusion models. The FSCDiff framework consists of two key components: the Dual-Domain Entanglement Enhancement Block (DTEB) and the Stable Time-step Mask Prediction Module (STMP). DTEB utilizes Fourier-spatial entanglement learning to fully exploit the Fourier and spatial domain information of RGB images and depth maps, thereby optimizing feature representation. STMP takes advantage of the excellent multi-step iterative mechanism of diffusion models to enhance the accuracy and robustness of the segmentation results. Comprehensive experimental results indicate that our FSCDiff method outperforms the state-of-the-art approaches on the USOD10K and USOD datasets. The source code is available at: https://github.com/lgwplay/FSCDiff.
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Your agent calls
Lunesearch_papers
Free to start. No credit card required.
Terminal
Install the CLIlune papers get df1a0518-4a98-4fc5-80f6-c6d113371c9cCited by top-tier papers1
Ask how each one uses itRelated papers
- WaterDiffusion: Learning a Prior-involved Unrolling Diffusion for Joint Underwater Saliency Detection and Visual RestorationLaibin Chang, Yunke Wang, Longxiang Deng, Bo Du et al.AAAI 2025 · 12 citations
- Wavelet-based Fourier Information Interaction with Frequency Diffusion Adjustment for Underwater Image RestorationChen Zhao, Weiling Cai, Chenyu Dong, Chengwei HuCVPR 2024 · 116 citations
- DiMSOD: A Diffusion-Based Framework for Multi-Modal Salient Object DetectionShuo Zhang, Jiaming Huang, Wenbing Tang, Yan Wu et al.AAAI 2025 · 3 citations
- SimpleDiffusion: A Lightweight and Efficient Conditional Diffusion Model for Multi-Modal Salient Object DetectionShuo Zhang, Jiaming Huang, Wenbing Tang, Jing Liu et al.AAAI 2026
- AMSP-UOD: When Vortex Convolution and Stochastic Perturbation Meet Underwater Object DetectionJingchun Zhou, Zongxin He, Kin-Man Lam, Yudong Wang et al.AAAI 2024 · 45 citations
