PROMOTE: Prior-Guided Diffusion Model with Global-Local Contrastive Learning for Exemplar-Based Image Translation
Guojin Zhong, Yihu Guo, Jin Yuan, Qianjun Zhang, Weili Guan, Long Chen
Abstract
Exemplar-based image translation has garnered significant interest from researchers due to its broad applications in multimedia/multimodal processing. Existing methods primarily employ Euclidean-based losses to implicitly establish cross-domain correspondences between exemplar and conditional images, aiming to produce high-fidelity images. However, these methods often suffer from two challenges: 1) Insufficient excavation of domain-invariant features leads to low-quality cross-domain correspondences, and 2) Inaccurate correspondences result in errors propagated during the translation process due to a lack of reliable prior guidance. To tackle these issues, we propose a novel prior-guided diffusion model with global-local contrastive learning (PROMOTE), which is trained in a self-supervised manner. Technically, global-local contrastive learning is designed to align two cross-domain images within hyperbolic space and reduce the gap between their semantic correlation distributions using the Fisher-Rao metric, allowing the visual encoders to extract domain-invariant features more effectively. Moreover, a prior-guided diffusion model is developed that propagates the structural prior to all timesteps in the diffusion process. It is optimized by a novel prior denoising loss, mathematically derived from the transitions modified by prior information in a self-supervised manner, successfully alleviating the impact of inaccurate correspondences on image translation. Extensive experiments conducted across seven datasets demonstrate that our proposed PROMOTE significantly exceeds state-of-the-art performance in diverse exemplar-based image translation tasks. The source code is publicly available at http://github.com/zgj77/PROMOTE.
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Your agent calls
Lunesearch_papers
Free to start. No credit card required.
Terminal
Install the CLIlune papers get 9bc97b9f-62f2-4e0a-94a1-3c6cb4e70941Cited by top-tier papers2
- DECIDER: Difference-aware Contrastive Diffusion Model with Adversarial Perturbations for Image Change CaptioningGuojin Zhong, Jinhong Hu, Jiajun Chen, Jin Yuan et al.AAAI 2025 · 3 citations
- Multi-Resolution Decomposable Diffusion Model for Non-Stationary Time Series Anomaly DetectionGuojin Zhong, Pan Wang, Jin Yuan, Zhiyong Li et al.ICLR 2025
Related papers
- Marginal Contrastive Correspondence for Guided Image GenerationFangneng Zhan, Yingchen Yu, Rongliang Wu, Jiahui Zhang et al.CVPR 2022 · 38 citations
- Accept the Modality Gap: An Exploration in the Hyperbolic SpaceSameera Ramasinghe, Violetta Shevchenko, Gil Avraham, Thalaiyasingam AjanthanCVPR 2024 · 11 citations
- Masked and Adaptive Transformer for Exemplar Based Image TranslationChang Jiang, Fei Gao, Biao Ma, Yuhao Lin et al.CVPR 2023
- Cross-Domain Correspondence Learning for Exemplar-Based Image TranslationPan Zhang, Bo Zhang, Dong Chen, Lu Yuan et al.CVPR 2020
- Diffusion-based Image Translation with Label Guidance for Domain Adaptive Semantic SegmentationDuo Peng, Ping Hu, Qiuhong Ke, Jun LiuICCV 2023 · 42 citations
