StyleRes: Transforming the Residuals for Real Image Editing with StyleGAN
Hamza Pehlivan, Yusuf Dalva, Aysegul Dundar
摘要
We present a novel image inversion framework and a training pipeline to achieve high-fidelity image inversion with high-quality attribute editing. Inverting real images into StyleGAN's latent space is an extensively studied problem, yet the trade-off between the image reconstruction fidelity and image editing quality remains an open challenge. The low-rate latent spaces are limited in their expressiveness power for high-fidelity reconstruction. On the other hand, high-rate latent spaces result in degradation in editing quality. In this work, to achieve high-fidelity inversion, we learn residual features in higher latent codes that lower latent codes were not able to encode. This enables preserving image details in reconstruction. To achieve high-quality editing, we learn how to transform the residual features for adapting to manipulations in latent codes. We train the framework to extract residual features and transform them via a novel architecture pipeline and cycle consistency losses. We run extensive experiments and compare our method with state-of-the-art inversion methods. Qualitative metrics and visual comparisons show significant improvements. Code: https://github.com/hamzapehlivan/StyleRes
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper15
- Schedule Your Edit: A Simple yet Effective Diffusion Noise Schedule for Image EditingHaonan Lin, Yan Chen, Jiahao Wang, Wenbin An 等NeurIPS 2024 · 被引用 46 次
- Diverse Inpainting and Editing with GAN InversionAhmet Burak Yildirim, Hamza Pehlivan, Bahri Batuhan Bilecen, Aysegul DundarICCV 2023 · 被引用 35 次
- SDGAN: Disentangling Semantic Manipulation for Facial Attribute EditingWenmin Huang, Weiqi Luo, Jiwu Huang, Xiaochun CaoAAAI 2024 · 被引用 20 次
- WildFake: A Large-Scale and Hierarchical Dataset for AI-Generated Images DetectionYan Hong, Jianming Feng, Haoxing Chen, Jun Lan 等AAAI 2025 · 被引用 13 次
- Dual Encoder GAN Inversion for High-Fidelity 3D Head Reconstruction from Single ImagesBahri Batuhan Bilecen, Ahmet Berke Gökmen, Aysegul DundarNeurIPS 2024 · 被引用 11 次
它引用的顶会 Paper21
- StyleCLIP: Text-Driven Manipulation of StyleGAN ImageryOr Patashnik, Zongze Wu, Eli Shechtman, Daniel Cohen-Or 等ICCV 2021 · 被引用 1,437 次
- Image2StyleGAN: How to Embed Images Into the StyleGAN Latent Space?Rameen Abdal, Yipeng Qin, Peter WonkaICCV 2019 · 被引用 1,195 次
- GANSpace: Discovering Interpretable GAN ControlsErik Härkönen, Aaron Hertzmann, Jaakko Lehtinen, Sylvain ParisNeurIPS 2020 · 被引用 1,049 次
- Unsupervised Discovery of Interpretable Directions in the GAN Latent SpaceAndrey Voynov, Artem BabenkoICML 2020 · 被引用 459 次
- ReStyle: A Residual-Based StyleGAN Encoder via Iterative RefinementYuval Alaluf, Or Patashnik, Daniel Cohen-OrICCV 2021 · 被引用 377 次
相关 Paper
- Style Transformer for Image Inversion and EditingXueqi Hu, Qiusheng Huang, Zhengyi Shi, Siyuan Li 等CVPR 2022 · 被引用 58 次
- ReGANIE: Rectifying GAN Inversion Errors for Accurate Real Image EditingBingchuan Li, Tianxiang Ma, Peng Zhang, Miao Hua 等AAAI 2023 · 被引用 11 次
- Designing an encoder for StyleGAN image manipulationOmer Tov, Yuval Alaluf, Yotam Nitzan, Or Patashnik 等SIGGRAPH 2021 · 被引用 692 次
- Self-Supervised Geometry-Aware Encoder for Style-Based 3D GAN InversionYushi Lan, Xuyi Meng, Shuai Yang, Chen Change Loy 等CVPR 2023
- A Latent Transformer for Disentangled Face Editing in Images and VideosXu Yao, Alasdair Newson, Yann Gousseau, Pierre HellierICCV 2021 · 被引用 97 次
