Do GANs always have Nash equilibria?
Farzan Farnia, Asuman E. Ozdaglar
摘要
Generative adversarial networks (GANs) represent a zero-sum game between two machine players, a generator and a discriminator, designed to learn the distribution of data. While GANs have achieved state-of-the-art performance in several benchmark learning tasks, GAN minimax optimization still poses great theoretical and empirical challenges. GANs trained using first-order optimization methods commonly fail to converge to a stable solution where the players cannot improve their objective, i.e., the Nash equilibrium of the underlying game. Such issues raise the question of the existence of Nash equilibria in GAN zero-sum games. In this work, we show through theoretical and numerical results that indeed GAN zero-sum games may have no Nash equilibria. To characterize an equilibrium notion applicable to GANs, we consider the equilibrium of a new zerosum game with an objective function given by a proximal operator applied to the original objective, a solution we call the proximal equilibrium. Unlike the Nash equilibrium, the proximal equilibrium captures the sequential nature of GANs, in which the generator moves first followed by the discriminator. We prove that the optimal generative model in Wasserstein GAN problems provides a proximal equilibrium. Inspired by these results, we propose a new approach, which we call proximal training, for solving GAN problems. We perform several numerical experiments indicating the existence of proximal equilibria in GANs.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper17
- If Influence Functions are the Answer, Then What is the Question?Juhan Bae, Nathan Ng, Alston Lo, Marzyeh Ghassemi 等NeurIPS 2022 · 被引用 185 次
- SAN: Inducing Metrizability of GAN with Discriminative Normalized Linear LayerYuhta Takida, Masaaki Imaizumi, Takashi Shibuya, Chieh-Hsin Lai 等ICLR 2024 · 被引用 28 次
- Hidden Convexity of Wasserstein GANs: Interpretable Generative Models with Closed-Form SolutionsArda Sahiner, Tolga Ergen, Batu Ozturkler, Burak Bartan 等ICLR 2022 · 被引用 23 次
- Free Lunch for Domain Adversarial Training: Environment Label SmoothingYifan Zhang, Xue Wang, Jian Liang, Zhang Zhang 等ICLR 2023 · 被引用 23 次
- On Convergence of Gradient Descent Ascent: A Tight Local AnalysisHaochuan Li, Farzan Farnia, Subhro Das, Ali JadbabaieICML 2022 · 被引用 12 次
它引用的顶会 Paper5
- On Gradient Descent Ascent for Nonconvex-Concave Minimax ProblemsTianyi Lin, Chi Jin, Michael I. JordanICML 2020 · 被引用 587 次
- On Solving Minimax Optimization Locally: A Follow-the-Ridge ApproachYuanhao Wang, Guodong Zhang, Jimmy BaICLR 2020 · 被引用 106 次
- A Closer Look at the Optimization Landscapes of Generative Adversarial NetworksHugo Berard, Gauthier Gidel, Amjad Almahairi, Pascal Vincent 等ICLR 2020 · 被引用 66 次
- SGD Learns One-Layer Networks in WGANsQi Lei, Jason D. Lee, Alex Dimakis, Constantinos DaskalakisICML 2020 · 被引用 36 次
- Implicit competitive regularization in GANsFlorian Schäfer, Hongkai Zheng, Animashree AnandkumarICML 2020 · 被引用 35 次
相关 Paper
- On Characterizing GAN Convergence Through Proximal Duality GapSahil Sidheekh, Aroof Aimen, Narayanan C. KrishnanICML 2021 · 被引用 7 次
- Provably convergent quasistatic dynamics for mean-field two-player zero-sum gamesChao Ma, Lexing YingICLR 2022 · 被引用 15 次
- A Convergent and Dimension-Independent Min-Max Optimization AlgorithmVijay Keswani, Oren Mangoubi, Sushant Sachdeva, Nisheeth K. VishnoiICML 2022 · 被引用 2 次
- DO-GAN: A Double Oracle Framework for Generative Adversarial NetworksAye Phyu Phyu Aung, Xinrun Wang, Runsheng Yu, Bo An 等CVPR 2022 · 被引用 1 次
- A Decentralized Parallel Algorithm for Training Generative Adversarial NetsMingrui Liu, Wei Zhang, Youssef Mroueh, Xiaodong Cui 等NeurIPS 2020 · 被引用 6 次
