Saliency-Guided Image Translation
Lai Jiang, Mai Xu, Xiaofei Wang, Leonid Sigal
Abstract
In this paper, we propose a novel task for saliencyguided image translation, with the goal of image-to-image translation conditioned on the user specified saliency map. To address this problem, we develop a novel Generative Adversarial Network (GAN)-based model, called SalG-GAN. Given the original image and target saliency map, SalG-GAN can generate a translated image that satisfies the target saliency map. In SalG-GAN, a disentangled representation framework is proposed to encourage the model to learn diverse translations for the same target saliency condition. A saliency-based attention module is introduced as a special attention mechanism for facilitating the developed structures of saliency-guided generator, saliency cue encoder and saliency-guided global and local discriminators. Furthermore, we build a synthetic dataset and a real-world dataset with labeled visual attention for training and evaluating our SalG-GAN. The experimental results over both datasets verify the effectiveness of our model for saliencyguided image translation.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 8230b242-592e-4b06-8c7f-9dcf4aca72eeCited by top-tier papers6
- Synthetic Data Supervised Salient Object DetectionZhenyu Wu, Lin Wang, Wei Wang, Tengfei Shi et al.ACM MM 2022 · 29 citations
- Target Scanpath-Guided 360-Degree Image EnhancementYujia Wang, Fang-Lue Zhang, Neil A. DodgsonAAAI 2025 · 21 citations
- Deep Saliency Prior for Reducing Visual DistractionKfir Aberman, Junfeng He, Yossi Gandelsman, Inbar Mosseri et al.CVPR 2022 · 20 citations
- Penetration Vision through Virtual Reality Headsets: Identifying 360-degree Videos from Head MovementsAnh Nguyen, Xiaokuan Zhang, Zhisheng YanUSENIX Security 2024 · 17 citations
- Sketch2Saliency: Learning to Detect Salient Objects from Human DrawingsAyan Kumar Bhunia, Subhadeep Koley, Amandeep Kumar, Aneeshan Sain et al.CVPR 2023
Builds on2
- Free-Form Image Inpainting With Gated ConvolutionJiahui Yu, Zhe Lin, Jimei Yang, Xiaohui Shen et al.ICCV 2019 · 1,990 citations
- Tell, Draw, and Repeat: Generating and Modifying Images Based on Continual Linguistic InstructionAlaaeldin El-Nouby, Shikhar Sharma, Hannes Schulz, R. Devon Hjelm et al.ICCV 2019 · 128 citations
Related papers
- Diagonal Attention and Style-based GAN for Content-Style Disentanglement in Image Generation and TranslationGihyun Kwon, Jong Chul YeICCV 2021 · 59 citations
- Unpaired Image Enhancement with Quality-Attention Generative Adversarial NetworkZhangkai Ni, Wenhan Yang, Shiqi Wang, Lin Ma et al.ACM MM 2020 · 23 citations
- DeSmoothGAN: Recovering Details of Smoothed Images via Spatial Feature-wise Transformation and Full AttentionYifei Huang, Chenhui Li, Xiaohu Guo, Jing Liao et al.ACM MM 2020 · 4 citations
- Semantics-Enhanced Adversarial Nets for Text-to-Image SynthesisHongchen Tan, Xiuping Liu, Xin Li, Yi Zhang et al.ICCV 2019 · 80 citations
- Style-Guided and Disentangled Representation for Robust Image-to-Image TranslationJaewoong Choi, Dae Ha Kim, Byung Cheol SongAAAI 2022 · 9 citations
