SAN: Inducing Metrizability of GAN with Discriminative Normalized Linear Layer
Yuhta Takida, Masaaki Imaizumi, Takashi Shibuya, Chieh-Hsin Lai, Toshimitsu Uesaka, Naoki Murata, Yuki Mitsufuji
摘要
Generative adversarial networks (GANs) learn a target probability distribution by optimizing a generator and a discriminator with minimax objectives. This paper addresses the question of whether such optimization actually provides the generator with gradients that make its distribution close to the target distribution. We derive metrizable conditions, sufficient conditions for the discriminator to serve as the distance between the distributions, by connecting the GAN formulation with the concept of sliced optimal transport. Furthermore, by leveraging these theoretical results, we propose a novel GAN training scheme called the Slicing Adversarial Network (SAN). With only simple modifications, a broad class of existing GANs can be converted to SANs. Experiments on synthetic and image datasets support our theoretical results and the effectiveness of SAN as compared to the usual GANs. We also apply SAN to StyleGAN-XL, which leads to a stateof-the-art FID score amongst GANs for class conditional generation on CIFAR10 and ImageNet 256×256. Our implementation is available on the project page https://ytakida.github.io/san/ .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- The GAN is dead; long live the GAN! A Modern GAN BaselineNick Huang, Aaron Gokaslan, Volodymyr Kuleshov, James TompkinNeurIPS 2024 · 被引用 111 次
- Neon: Negative Extrapolation From Self-Training Improves Image GenerationSina Alemohammad, Zhangyang Wang, Richard BaraniukICLR 2026 · 被引用 5 次
- Community Forensics: Using Thousands of Generators to Train Fake Image DetectorsJeongsoo Park, Andrew OwensCVPR 2025
- SONA: Learning Conditional, Unconditional, and Matching-Aware DiscriminatorYuhta Takida, Satoshi Hayakawa, Takashi Shibuya, Masaaki Imaizumi 等ICLR 2026
它引用的顶会 Paper30
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 被引用 11,743 次
- Training data-efficient image transformers & distillation through attentionHugo Touvron, Matthieu Cord, Matthijs Douze, Francisco Massa 等ICML 2021 · 被引用 8,974 次
- Improved Denoising Diffusion Probabilistic ModelsAlexander Quinn Nichol, Prafulla DhariwalICML 2021 · 被引用 5,234 次
相关 Paper
- Bridging the Gap Between f-GANs and Wasserstein GANsJiaming Song, Stefano ErmonICML 2020 · 被引用 45 次
- Run-Sort-ReRun: Escaping Batch Size Limitations in Sliced Wasserstein Generative ModelsJosé Lezama, Wei Chen, Qiang QiuICML 2021 · 被引用 9 次
- A Unified View of cGANs with and without ClassifiersSi-An Chen, Chun-Liang Li, Hsuan-Tien LinNeurIPS 2021 · 被引用 12 次
- Spider GAN: Leveraging Friendly Neighbors to Accelerate GAN TrainingSiddarth Asokan, Chandra Sekhar SeelamantulaCVPR 2023
- Wasserstein-2 Generative NetworksAlexander Korotin, Vage Egiazarian, Arip Asadulaev, Alexander Safin 等ICLR 2021 · 被引用 128 次
