Polymorphic-GAN: Generating Aligned Samples across Multiple Domains with Learned Morph Maps
Seung Wook Kim, Karsten Kreis, Daiqing Li, Antonio Torralba, Sanja Fidler
摘要
Modern image generative models show remarkable sample quality when trained on a single domain or class of objects. In this work, we introduce a generative adversarial network that can simultaneously generate aligned image samples from multiple related domains. We leverage the fact that a variety of object classes share common attributes, with certain geometric differences. We propose Polymorphic-GAN which learns shared features across all domains and a per-domain morph layer to morph shared features according to each domain. In contrast to previous works, our framework allows simultaneous modelling of images with highly varying geometries, such as images of human faces, painted and artistic faces, as well as multiple different animal faces. We demonstrate that our model produces aligned samples for all domains and show how it can be used for applications such as segmentation transfer and cross-domain image editing, as well as training in low-data regimes. Additionally, we apply our Polymorphic-GAN on image-to-image translation tasks and show that we can greatly surpass previous approaches in cases where the geometric differences between domains are large.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- NeuralField-LDM: Scene Generation with Hierarchical Latent Diffusion ModelsSeung Wook Kim, Bradley Brown, Kangxue Yin, Karsten Kreis 等CVPR 2023
- Probability Density Geodesics in Image Diffusion Latent SpaceQingtao Yu, Jaskirat Singh, Zhaoyuan Yang, Peter Henry Tu 等CVPR 2025
它引用的顶会 Paper28
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- Training Generative Adversarial Networks with Limited DataTero Karras, Miika Aittala, Janne Hellsten, Samuli Laine 等NeurIPS 2020 · 被引用 2,345 次
- StyleCLIP: Text-Driven Manipulation of StyleGAN ImageryOr Patashnik, Zongze Wu, Eli Shechtman, Daniel Cohen-Or 等ICCV 2021 · 被引用 1,437 次
- Image2StyleGAN: How to Embed Images Into the StyleGAN Latent Space?Rameen Abdal, Yipeng Qin, Peter WonkaICCV 2019 · 被引用 1,195 次
- GANSpace: Discovering Interpretable GAN ControlsErik Härkönen, Aaron Hertzmann, Jaakko Lehtinen, Sylvain ParisNeurIPS 2020 · 被引用 1,049 次
相关 Paper
- StyleAlign: Analysis and Applications of Aligned StyleGAN ModelsZongze Wu, Yotam Nitzan, Eli Shechtman, Dani LischinskiICLR 2022 · 被引用 65 次
- DeepI2I: Enabling Deep Hierarchical Image-to-Image Translation by Transferring from GANsYaxing Wang, Lu Yu, Joost van de WeijerNeurIPS 2020 · 被引用 18 次
- RelGAN: Multi-Domain Image-to-Image Translation via Relative AttributesYu-Jing Lin, Po-Wei Wu, Che-Han Chang, Edward Y. Chang 等ICCV 2019 · 被引用 158 次
- 3DAvatarGAN: Bridging Domains for Personalized Editable AvatarsRameen Abdal, Hsin-Ying Lee, Peihao Zhu, Menglei Chai 等CVPR 2023
- Attribute Manipulation Generative Adversarial Networks for Fashion ImagesKenan E. Ak, Ashraf A. Kassim, Joo-Hwee Lim, Jo Yew ThamICCV 2019 · 被引用 85 次
