On Characterizing GAN Convergence Through Proximal Duality Gap
Sahil Sidheekh, Aroof Aimen, Narayanan C. Krishnan
摘要
Despite the accomplishments of Generative Adversarial Networks (GANs) in modeling data distributions, training them remains a challenging task. A contributing factor to this difficulty is the non-intuitive nature of the GAN loss curves, which necessitates a subjective evaluation of the generated output to infer training progress. Recently, motivated by game theory, duality gap has been proposed as a domain agnostic measure to monitor GAN training. However, it is restricted to the setting when the GAN converges to a Nash equilibrium. But GANs need not always converge to a Nash equilibrium to model the data distribution. In this work, we extend the notion of duality gap to proximal duality gap that is applicable to the general context of training GANs where Nash equilibria may not exist. We show theoretically that the proximal duality gap is capable of monitoring the convergence of GANs to a wider spectrum of equilibria that subsumes Nash equilibria. We also theoretically establish the relationship between the proximal duality gap and the divergence between the real and generated data distributions for different GAN formulations. Our results provide new insights into the nature of GAN convergence. Finally, we validate experimentally the usefulness of proximal duality gap for monitoring and influencing GAN training.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper5
- On Gradient Descent Ascent for Nonconvex-Concave Minimax ProblemsTianyi Lin, Chi Jin, Michael I. JordanICML 2020 · 被引用 587 次
- Reliable Fidelity and Diversity Metrics for Generative ModelsMuhammad Ferjad Naeem, Seong Joon Oh, Youngjung Uh, Yunjey Choi 等ICML 2020 · 被引用 553 次
- What is Local Optimality in Nonconvex-Nonconcave Minimax Optimization?Chi Jin, Praneeth Netrapalli, Michael I. JordanICML 2020 · 被引用 381 次
- Do GANs always have Nash equilibria?Farzan Farnia, Asuman E. OzdaglarICML 2020 · 被引用 93 次
- A Closer Look at the Optimization Landscapes of Generative Adversarial NetworksHugo Berard, Gauthier Gidel, Amjad Almahairi, Pascal Vincent 等ICLR 2020 · 被引用 66 次
相关 Paper
- MonoFlow: Rethinking Divergence GANs via the Perspective of Wasserstein Gradient FlowsMingxuan Yi, Zhanxing Zhu, Song LiuICML 2023 · 被引用 18 次
- Understanding and Stabilizing GANs' Training Dynamics Using Control TheoryKun Xu, Chongxuan Li, Jun Zhu, Bo ZhangICML 2020 · 被引用 31 次
- DO-GAN: A Double Oracle Framework for Generative Adversarial NetworksAye Phyu Phyu Aung, Xinrun Wang, Runsheng Yu, Bo An 等CVPR 2022 · 被引用 1 次
- A Neural Tangent Kernel Perspective of GANsJean-Yves Franceschi, Emmanuel de Bézenac, Ibrahim Ayed, Mickaël Chen 等ICML 2022 · 被引用 29 次
- Inconsistency, Instability, and Generalization Gap of Deep Neural Network TrainingRie Johnson, Tong ZhangNeurIPS 2023 · 被引用 11 次
