DUNIT: Detection-Based Unsupervised Image-to-Image Translation
Deblina Bhattacharjee, Seungryong Kim, Guillaume Vizier, Mathieu Salzmann
摘要
Image-to-image translation has made great strides in recent years, with current techniques being able to handle unpaired training images and to account for the multimodality of the translation problem. Despite this, most methods treat the image as a whole, which makes the results they produce for content-rich scenes less realistic. In this paper, we introduce a Detection-based Unsupervised Image-to-image Translation (DUNIT) approach that explicitly accounts for the object instances in the translation process. To this end, we extract separate representations for the global image and for the instances, which we then fuse into a common representation from which we generate the translated image. This allows us to preserve the detailed content of object instances, while still modeling the fact that we aim to produce an image of a single consistent scene. We introduce an instance consistency loss to maintain the coherence between the detections. Furthermore, by incorporating a detector into our architecture, we can still exploit object instances at test time. As evidenced by our experiments, this allows us to outperform the state-of-the-art unsupervised image-to-image translation methods. Furthermore, our approach can also be used as an unsupervised domain adaptation strategy for object detection, and it also achieves state-of-the-art performance on this task.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper15
- Frequency Domain Image Translation: More Photo-realistic, Better Identity-preservingMu Cai, Hong Zhang, Huijuan Huang, Qichuan Geng 等ICCV 2021 · 被引用 118 次
- Dual Path Learning for Domain Adaptation of Semantic SegmentationYiting Cheng, Fangyun Wei, Jianmin Bao, Dong Chen 等ICCV 2021 · 被引用 77 次
- MuIT: An End-to-End Multitask Learning TransformerDeblina Bhattacharjee, Tong Zhang, Sabine Süsstrunk, Mathieu SalzmannCVPR 2022 · 被引用 64 次
- InstaFormer: Instance-Aware Image-to-Image Translation with TransformerSoohyun Kim, Jongbeom Baek, Jihye Park, Gyeongnyeon Kim 等CVPR 2022 · 被引用 53 次
- Few shot font generation via transferring similarity guided global style and quantization local styleWei Pan, Anna Zhu, Xinyu Zhou, Brian Kenji Iwana 等ICCV 2023 · 被引用 24 次
相关 Paper
- Rethinking the Truly Unsupervised Image-to-Image TranslationKyungjune Baek, Yunjey Choi, Youngjung Uh, Jaejun Yoo 等ICCV 2021 · 被引用 115 次
- Self-Supervised Dense Consistency Regularization for Image-to-Image TranslationMinsu Ko, Eunju Cha, Sungjoo Suh, Huijin Lee 等CVPR 2022 · 被引用 25 次
- Cross-Granularity Learning for Multi-Domain Image-to-Image TranslationHuiyuan Fu, Ting Yu, Xin Wang, Huadong MaACM MM 2020 · 被引用 4 次
- Memory-Guided Unsupervised Image-to-Image TranslationSomi Jeong, Youngjung Kim, Eungbean Lee, Kwanghoon SohnCVPR 2021
- Unsupervised Image-to-Image Translation with Generative PriorShuai Yang, Liming Jiang, Ziwei Liu, Chen Change LoyCVPR 2022 · 被引用 51 次
