Energy-Based Generative Cooperative Saliency Prediction
Jing Zhang, Jianwen Xie, Zilong Zheng, Nick Barnes
摘要
Conventional saliency prediction models typically learn a deterministic mapping from an image to its saliency map, and thus fail to explain the subjective nature of human attention. In this paper, to model the uncertainty of visual saliency, we study the saliency prediction problem from the perspective of generative models by learning a conditional probability distribution over the saliency map given an input image, and treating the saliency prediction as a sampling process from the learned distribution. Specifically, we propose a generative cooperative saliency prediction framework, where a conditional latent variable model (LVM) and a conditional energy-based model (EBM) are jointly trained to predict salient objects in a cooperative manner. The LVM serves as a fast but coarse predictor to efficiently produce an initial saliency map, which is then refined by the iterative Langevin revision of the EBM that serves as a slow but fine predictor. Such a coarse-to-fine cooperative saliency prediction strategy offers the best of both worlds. Moreover, we propose a ``cooperative learning while recovering" strategy and apply it to weakly supervised saliency prediction, where saliency annotations of training images are partially observed. Lastly, we find that the learned energy function in the EBM can serve as a refinement module that can refine the results of other pre-trained saliency prediction models. Experimental results show that our model can produce a set of diverse and plausible saliency maps of an image, and obtain state-of-the-art performance in both fully supervised and weakly supervised saliency prediction tasks.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- Learning Generative Vision Transformer with Energy-Based Latent Space for Saliency PredictionJing Zhang, Jianwen Xie, Nick Barnes, Ping LiNeurIPS 2021 · 被引用 117 次
- Improving Adversarial Energy-Based Model via Diffusion ProcessCong Geng, Tian Han, Peng-Tao Jiang, Hao Zhang 等ICML 2024 · 被引用 5 次
- CoopInit: Initializing Generative Adversarial Networks via Cooperative LearningYang Zhao, Jianwen Xie, Ping LiAAAI 2023 · 被引用 3 次
它引用的顶会 Paper11
- Your classifier is secretly an energy based model and you should treat it like oneWill Grathwohl, Kuan-Chieh Wang, Jörn-Henrik Jacobsen, David Duvenaud 等ICLR 2020 · 被引用 643 次
- Stacked Cross Refinement Network for Edge-Aware Salient Object DetectionZhe Wu, Li Su, Qingming HuangICCV 2019 · 被引用 374 次
- Locate Globally, Segment Locally: A Progressive Architecture With Knowledge Review Network for Salient Object DetectionBinwei Xu, Haoran Liang, Ronghua Liang, Peng ChenAAAI 2021 · 被引用 186 次
- Structure-Consistent Weakly Supervised Salient Object Detection with Local Saliency CoherenceSiyue Yu, Bingfeng Zhang, Jimin Xiao, Eng Gee LimAAAI 2021 · 被引用 162 次
- Learning Energy-Based Model with Variational Auto-Encoder as Amortized SamplerJianwen Xie, Zilong Zheng, Ping LiAAAI 2021 · 被引用 57 次
相关 Paper
- Saliency-Guided Image TranslationLai Jiang, Mai Xu, Xiaofei Wang, Leonid SigalCVPR 2021
- CSDN: CLIP-Driven Similarity-Aligned Distillation Network for Weakly-Supervised Object LocalizationSifan Zuo, Youfa Liu, Bo DuACM MM 2025
- Learning Energy-Based Generative Models via Coarse-to-Fine Expanding and SamplingYang Zhao, Jianwen Xie, Ping LiICLR 2021 · 被引用 51 次
- Railroad Is Not a Train: Saliency As Pseudo-Pixel Supervision for Weakly Supervised Semantic SegmentationSeungho Lee, Minhyun Lee, Jongwuk Lee, Hyunjung ShimCVPR 2021
- Sketch2Saliency: Learning to Detect Salient Objects from Human DrawingsAyan Kumar Bhunia, Subhadeep Koley, Amandeep Kumar, Aneeshan Sain 等CVPR 2023
