Few-shot Cross-domain Image Generation via Inference-time Latent-code Learning
Arnab Kumar Mondal, Piyush Tiwary, Parag Singla, Prathosh AP
摘要
In this work, our objective is to adapt a Deep generative model trained on a largescale source dataset to multiple target domains with scarce data. Specifically, we focus on adapting a pre-trained Generative Adversarial Network (GAN) to a target domain without re-training the generator. Our method draws the motivation from the fact that out-of-distribution samples can be 'embedded' onto the latent space of a pre-trained source-GAN. We propose to train a small latent-generation network during the inference stage, each time a batch of target samples is to be generated. These target latent codes are fed to the source-generator to obtain novel target samples. Despite using the same small set of target samples and the source generator, multiple independent training episodes of the latent-generation network results in the diversity of the generated target samples. Our method, albeit simple, can be used to generate data from multiple target distributions using a generator trained on a single source distribution. We demonstrate the efficacy of our surprisingly simple method in generating multiple target datasets with only a single source generator and a few target samples. The code of the proposed method is available at: https://github.com/arnabkmondal/GenDA RELATED WORK Few shot generative domain adaptation: In 'generative domain adaptation', a base model pretrained on source domain is adapted to a related target domain by using few examples. Generally, this is done by re-training the model on the target data via appropriate losses. For example, the authors of Transfer-GAN (Wang et al., 2018) demonstrated that fine-tuning from a single pretrained GAN (Goodfellow et al., 2014) is beneficial for domains with scarce data. Later, the authors in (Noguchi & Harada, 2019) observed that this technique leads to mode collapse, and hence they only fine-tune the scale and shift parameters of the generator. However, this may limit the flexibility of the network. To address this concern, the authors in MineGAN (Wang et al., 2020b) prepend a miner network to the generator to transform the input latent space modeled by multivariate normal distribution so that the generated images resemble the target domain. They propose a two step-training
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- Deformable One-Shot Face Stylization via DINO Semantic GuidanceYang Zhou, Zichong Chen, Hui HuangCVPR 2024 · 被引用 9 次
- Few-shot Hybrid Domain Adaptation of Image GeneratorHengjia Li, Yang Liu, Linxuan Xia, Yuqi Lin 等ICLR 2024 · 被引用 7 次
- Towards Scalable Topological RegularizersHiu-Tung Wong, Darrick Lee, Hong YanICLR 2025
它引用的顶会 Paper19
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- Training Generative Adversarial Networks with Limited DataTero Karras, Miika Aittala, Janne Hellsten, Samuli Laine 等NeurIPS 2020 · 被引用 2,345 次
- Image2StyleGAN: How to Embed Images Into the StyleGAN Latent Space?Rameen Abdal, Yipeng Qin, Peter WonkaICCV 2019 · 被引用 1,195 次
- Differentiable Augmentation for Data-Efficient GAN TrainingShengyu Zhao, Zhijian Liu, Ji Lin, Jun-Yan Zhu 等NeurIPS 2020 · 被引用 707 次
- Designing an encoder for StyleGAN image manipulationOmer Tov, Yuval Alaluf, Yotam Nitzan, Or Patashnik 等SIGGRAPH 2021 · 被引用 692 次
相关 Paper
- One-Shot Generative Domain AdaptationCeyuan Yang, Yujun Shen, Zhiyi Zhang, Yinghao Xu 等ICCV 2023 · 被引用 69 次
- Few-shot Image Generation with Elastic Weight ConsolidationYijun Li, Richard Zhang, Jingwan Lu, Eli ShechtmanNeurIPS 2020 · 被引用 193 次
- MineGAN: Effective Knowledge Transfer From GANs to Target Domains With Few ImagesYaxing Wang, Abel Gonzalez-Garcia, David Berga, Luis Herranz 等CVPR 2020
- Image Generation From Small Datasets via Batch Statistics AdaptationAtsuhiro Noguchi, Tatsuya HaradaICCV 2019 · 被引用 211 次
- WeditGAN: Few-Shot Image Generation via Latent Space RelocationYuxuan Duan, Li Niu, Yan Hong, Liqing ZhangAAAI 2024 · 被引用 23 次
