Enjoy Your Editing: Controllable GANs for Image Editing via Latent Space Navigation
Peiye Zhuang, Oluwasanmi Koyejo, Alexander G. Schwing
摘要
Controllable semantic image editing enables a user to change entire image attributes with few clicks, e.g., gradually making a summer scene look like it was taken in winter. Classic approaches for this task use a Generative Adversarial Net (GAN) to learn a latent space and suitable latent-space transformations. However, current approaches often suffer from attribute edits that are entangled, global image identity changes, and diminished photo-realism. To address these concerns, we learn multiple attribute transformations simultaneously, we integrate attribute regression into the training of transformation functions, apply a content loss and an adversarial loss that encourage the maintenance of image identity and photo-realism. We propose quantitative evaluation strategies for measuring controllable editing performance, unlike prior work which primarily focuses on qualitative evaluation. Our model permits better control for both single- and multiple-attribute editing, while also preserving image identity and realism during transformation. We provide empirical results for both real and synthetic images, highlighting that our model achieves state-of-the-art performance for targeted image manipulation.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper23
- High-Fidelity GAN Inversion for Image Attribute EditingTengfei Wang, Yong Zhang, Yanbo Fan, Jue Wang 等CVPR 2022 · 被引用 227 次
- Talk-to-Edit: Fine-Grained Facial Editing via DialogYuming Jiang, Ziqi Huang, Xingang Pan, Chen Change Loy 等ICCV 2021 · 被引用 162 次
- GAN-Control: Explicitly Controllable GANsAlon Shoshan, Nadav Bhonker, Igor Kviatkovsky, Gérard G. MedioniICCV 2021 · 被引用 151 次
- StyleGAN knows Normal, Depth, Albedo, and MoreAnand Bhattad, Daniel McKee, Derek Hoiem, David A. ForsythNeurIPS 2023 · 被引用 61 次
- Schedule Your Edit: A Simple yet Effective Diffusion Noise Schedule for Image EditingHaonan Lin, Yan Chen, Jiahao Wang, Wenbin An 等NeurIPS 2024 · 被引用 46 次
它引用的顶会 Paper9
- Image2StyleGAN: How to Embed Images Into the StyleGAN Latent Space?Rameen Abdal, Yipeng Qin, Peter WonkaICCV 2019 · 被引用 1,195 次
- Unsupervised Discovery of Interpretable Directions in the GAN Latent SpaceAndrey Voynov, Artem BabenkoICML 2020 · 被引用 459 次
- On the "steerability" of generative adversarial networksAli Jahanian, Lucy Chai, Phillip IsolaICLR 2020 · 被引用 421 次
- RelGAN: Multi-Domain Image-to-Image Translation via Relative AttributesYu-Jing Lin, Po-Wei Wu, Che-Han Chang, Edward Y. Chang 等ICCV 2019 · 被引用 158 次
- Detecting Photoshopped Faces by Scripting PhotoshopSheng-Yu Wang, Oliver Wang, Richard Zhang, Andrew Owens 等ICCV 2019 · 被引用 147 次
相关 Paper
- SSFlow: Style-guided Neural Spline Flows for Face Image ManipulationHanbang Liang, Xianxu Hou, Linlin ShenACM MM 2021 · 被引用 11 次
- Adaptive Nonlinear Latent Transformation for Conditional Face EditingZhizhong Huang, Siteng Ma, Junping Zhang, Hongming ShanICCV 2023 · 被引用 13 次
- A Latent Transformer for Disentangled Face Editing in Images and VideosXu Yao, Alasdair Newson, Yann Gousseau, Pierre HellierICCV 2021 · 被引用 97 次
- Everything is There in Latent Space: Attribute Editing and Attribute Style Manipulation by StyleGAN Latent Space ExplorationRishubh Parihar, Ankit Dhiman, Tejan Karmali, Venkatesh Babu R.ACM MM 2022 · 被引用 21 次
- Text-Guided Unsupervised Latent Transformation for Multi-Attribute Image ManipulationXiwen Wei, Zhen Xu, Cheng Liu, Si Wu 等CVPR 2023
