On Characterizing GAN Convergence Through Proximal Duality Gap
Sahil Sidheekh, Aroof Aimen, Narayanan C. Krishnan
Abstract
Despite the accomplishments of Generative Adversarial Networks (GANs) in modeling data distributions, training them remains a challenging task. A contributing factor to this difficulty is the non-intuitive nature of the GAN loss curves, which necessitates a subjective evaluation of the generated output to infer training progress. Recently, motivated by game theory, duality gap has been proposed as a domain agnostic measure to monitor GAN training. However, it is restricted to the setting when the GAN converges to a Nash equilibrium. But GANs need not always converge to a Nash equilibrium to model the data distribution. In this work, we extend the notion of duality gap to proximal duality gap that is applicable to the general context of training GANs where Nash equilibria may not exist. We show theoretically that the proximal duality gap is capable of monitoring the convergence of GANs to a wider spectrum of equilibria that subsumes Nash equilibria. We also theoretically establish the relationship between the proximal duality gap and the divergence between the real and generated data distributions for different GAN formulations. Our results provide new insights into the nature of GAN convergence. Finally, we validate experimentally the usefulness of proximal duality gap for monitoring and influencing GAN training.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 45031be0-0734-4d22-b5f5-e0c9a01803aeCited by top-tier papers1
Ask how each one uses itBuilds on5
- On Gradient Descent Ascent for Nonconvex-Concave Minimax ProblemsTianyi Lin, Chi Jin, Michael I. JordanICML 2020 · 587 citations
- Reliable Fidelity and Diversity Metrics for Generative ModelsMuhammad Ferjad Naeem, Seong Joon Oh, Youngjung Uh, Yunjey Choi et al.ICML 2020 · 553 citations
- What is Local Optimality in Nonconvex-Nonconcave Minimax Optimization?Chi Jin, Praneeth Netrapalli, Michael I. JordanICML 2020 · 381 citations
- Do GANs always have Nash equilibria?Farzan Farnia, Asuman E. OzdaglarICML 2020 · 93 citations
- A Closer Look at the Optimization Landscapes of Generative Adversarial NetworksHugo Berard, Gauthier Gidel, Amjad Almahairi, Pascal Vincent et al.ICLR 2020 · 66 citations
Related papers
- MonoFlow: Rethinking Divergence GANs via the Perspective of Wasserstein Gradient FlowsMingxuan Yi, Zhanxing Zhu, Song LiuICML 2023 · 18 citations
- Understanding and Stabilizing GANs' Training Dynamics Using Control TheoryKun Xu, Chongxuan Li, Jun Zhu, Bo ZhangICML 2020 · 31 citations
- DO-GAN: A Double Oracle Framework for Generative Adversarial NetworksAye Phyu Phyu Aung, Xinrun Wang, Runsheng Yu, Bo An et al.CVPR 2022 · 1 citation
- A Neural Tangent Kernel Perspective of GANsJean-Yves Franceschi, Emmanuel de Bézenac, Ibrahim Ayed, Mickaël Chen et al.ICML 2022 · 29 citations
- Inconsistency, Instability, and Generalization Gap of Deep Neural Network TrainingRie Johnson, Tong ZhangNeurIPS 2023 · 11 citations
