S3OD: Towards Generalizable Salient Object Detection with Synthetic Data
Orest Kupyn, Hirokatsu Kataoka, Christian Rupprecht
摘要
Salient object detection exemplifies data-bounded tasks where expensive pixel-precise annotations force separate model training for related subtasks like DIS and HR-SOD. We present a method that dramatically improves generalization through large-scale synthetic data generation and ambiguity-aware architecture. We introduce S3OD, a dataset of over 139,000 high-resolution images created through our multi-modal diffusion pipeline that extracts labels from diffusion and DINO-v3 features. The iterative generation framework prioritizes challenging categories based on model performance. We propose a streamlined multi-mask decoder that handles the inherent ambiguity in salient object detection by predicting multiple valid interpretations. Models trained only on synthetic data achieve 20-50% error reduction in cross-dataset generalization, while fine-tuned versions reach state-of-the-art performance across DIS and HR-SOD benchmarks.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper25
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu 等ICCV 2021 · 被引用 31,683 次
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Vision Transformers for Dense PredictionRené Ranftl, Alexey Bochkovskiy, Vladlen KoltunICCV 2021 · 被引用 2,647 次
- EGNet: Edge Guidance Network for Salient Object DetectionJiaxing Zhao, Jiang-Jiang Liu, Deng-Ping Fan, Yang Cao 等ICCV 2019 · 被引用 1,054 次
相关 Paper
- Synthetic Data Supervised Salient Object DetectionZhenyu Wu, Lin Wang, Wei Wang, Tengfei Shi 等ACM MM 2022 · 被引用 29 次
- ODGEN: Domain-specific Object Detection Data Generation with Diffusion ModelsJingyuan Zhu, Shiyu Li, Yuxuan Liu, Jian Yuan 等NeurIPS 2024 · 被引用 32 次
- Disentangled High Quality Salient Object DetectionLv Tang, Bo Li, Yijie Zhong, Shouhong Ding 等ICCV 2021 · 被引用 86 次
- Annotation Ambiguity Aware Semi-Supervised Medical Image SegmentationSuruchi Kumari, Pravendra SinghCVPR 2025
- DiMSOD: A Diffusion-Based Framework for Multi-Modal Salient Object DetectionShuo Zhang, Jiaming Huang, Wenbing Tang, Yan Wu 等AAAI 2025 · 被引用 3 次
