PuppetGAN: Cross-Domain Image Manipulation by Demonstration
Ben Usman, Nick Dufour, Kate Saenko, Chris Bregler
Abstract
In this work we propose a model that can manipulate individual visual attributes of objects in a real scene using examples of how respective attribute manipulations affect the output of a simulation. As an example, we train our model to manipulate the expression of a human face using nonphotorealistic 3D renders of a face with varied expression. Our model manages to preserve all other visual attributes of a real face, such as head orientation, even though this and other attributes are not labeled in either real or synthetic domain. Since our model learns to manipulate a specific property in isolation using only "synthetic demonstrations" of such manipulations without explicitly provided labels, it can be applied to shape, texture, lighting, and other properties that are difficult to measure or represent as real-valued vectors. We measure the degree to which our model preserves other attributes of a real image when a single specific attribute is manipulated. We use digit datasets to analyze how discrepancy in attribute distributions affects the performance of our model, and demonstrate results in a far more difficult setting: learning to manipulate real human faces using nonphotorealistic 3D renders. Supported in part by DARPA and NSF awards IIS-1724237, CNS-1629700, CCF-1723379 (b) At test time we want to perform this manipulation with synthetic references on real images. rotation (a) At train time we receive demonstrations of a manipulation in the synthetic domain and examples of real images.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 06c497a3-90c7-4a13-ae72-fff1b3a25569Cited by top-tier papers5
- Learning Attribute-driven Disentangled Representations for Interactive Fashion RetrievalYuxin Hou, Eleonora Vig, Michael Donoser, Loris BazzaniICCV 2021 · 58 citations
- MOST-GAN: 3D Morphable StyleGAN for Disentangled Face Image ManipulationSafa C. Medin, Bernhard Egger, Anoop Cherian, Ye Wang et al.AAAI 2022 · 38 citations
- Retrieve in Style: Unsupervised Facial Feature Transfer and RetrievalMin Jin Chong, Wen-Sheng Chu, Abhishek Kumar, David A. ForsythICCV 2021 · 27 citations
- Analogical Image Translation for Fog GenerationRui Gong, Dengxin Dai, Yuhua Chen, Wen Li et al.AAAI 2021 · 19 citations
- StyleRig: Rigging StyleGAN for 3D Control Over Portrait ImagesAyush Tewari, Mohamed A. Elgharib, Gaurav Bharaj, Florian Bernard et al.CVPR 2020
Related papers
- FaceController: Controllable Attribute Editing for Face in the WildZhiliang Xu, Xiyu Yu, Zhibin Hong, Zhen Zhu et al.AAAI 2021 · 49 citations
- PERSE: Personalized 3D Generative Avatars from A Single PortraitHyunsoo Cha, Inhee Lee, Hanbyul JooCVPR 2025
- Enjoy Your Editing: Controllable GANs for Image Editing via Latent Space NavigationPeiye Zhuang, Oluwasanmi Koyejo, Alexander G. SchwingICLR 2021 · 88 citations
- Quantitative Manipulation of Custom Attributes on 3D-Aware Image SynthesisHoseok Do, Eunkyung Yoo, Taehyeong Kim, Chul Lee et al.CVPR 2023
- GAN-Control: Explicitly Controllable GANsAlon Shoshan, Nadav Bhonker, Igor Kviatkovsky, Gérard G. MedioniICCV 2021 · 151 citations
