Pseudo-Spherical Contrastive Divergence
Lantao Yu, Jiaming Song, Yang Song, Stefano Ermon
摘要
Energy-based models (EBMs) offer flexible distribution parametrization. However, due to the intractable partition function, they are typically trained via contrastive divergence for maximum likelihood estimation. In this paper, we propose pseudo-spherical contrastive divergence (PS-CD) to generalize maximum likelihood learning of EBMs. PS-CD is derived from the maximization of a family of strictly proper homogeneous scoring rules, which avoids the computation of the intractable partition function and provides a generalized family of learning objectives that include contrastive divergence as a special case. Moreover, PS-CD allows us to flexibly choose various learning objectives to train EBMs without additional computational cost or variational minimax optimization. Theoretical analysis on the proposed method and extensive experiments on both synthetic data and commonly used image datasets demonstrate the effectiveness and modeling flexibility of PS-CD, as well as its robustness to data contamination, thus showing its superiority over maximum likelihood and -EBMs.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- Energy-Based Models for Anomaly Detection: A Manifold Diffusion Recovery ApproachSangwoong Yoon, Young-Uk Jin, Yung-Kyun Noh, Frank C. ParkNeurIPS 2023 · 被引用 28 次
- RényiCL: Contrastive Representation Learning with Skew Rényi DivergenceKyungmin Lee, Jinwoo ShinNeurIPS 2022 · 被引用 13 次
- Language Generation with Strictly Proper Scoring RulesChenze Shao, Fandong Meng, Yijin Liu, Jie ZhouICML 2024 · 被引用 7 次
- Guiding Energy-based Models via Contrastive Latent VariablesHankook Lee, Jongheon Jeong, Sejun Park, Jinwoo ShinICLR 2023 · 被引用 4 次
它引用的顶会 Paper8
- Improved Techniques for Training Score-Based Generative ModelsYang Song, Stefano ErmonNeurIPS 2020 · 被引用 1,527 次
- Your classifier is secretly an energy based model and you should treat it like oneWill Grathwohl, Kuan-Chieh Wang, Jörn-Henrik Jacobsen, David Duvenaud 等ICLR 2020 · 被引用 643 次
- Reliable Fidelity and Diversity Metrics for Generative ModelsMuhammad Ferjad Naeem, Seong Joon Oh, Youngjung Uh, Yunjey Choi 等ICML 2020 · 被引用 553 次
- No MCMC for me: Amortized sampling for fast and stable training of energy-based modelsWill Sussman Grathwohl, Jacob Jin Kelly, Milad Hashemi, Mohammad Norouzi 等ICLR 2021 · 被引用 75 次
- Training Deep Energy-Based Models with f-Divergence MinimizationLantao Yu, Yang Song, Jiaming Song, Stefano ErmonICML 2020 · 被引用 50 次
相关 Paper
- Energy Discrepancies: A Score-Independent Loss for Energy-Based ModelsTobias Schröder, Zijing Ou, Jen Lim, Yingzhen Li 等NeurIPS 2023 · 被引用 15 次
- Variational (Gradient) Estimate of the Score Function in Energy-based Latent Variable ModelsFan Bao, Kun Xu, Chongxuan Li, Lanqing Hong 等ICML 2021 · 被引用 10 次
- Learning Energy-Based Models by Diffusion Recovery LikelihoodRuiqi Gao, Yang Song, Ben Poole, Ying Nian Wu 等ICLR 2021 · 被引用 144 次
- Bi-level Score Matching for Learning Energy-based Latent Variable ModelsFan Bao, Chongxuan Li, Taufik Xu, Hang Su 等NeurIPS 2020 · 被引用 16 次
- Joint Learning of Energy-based Models and their Partition FunctionMichael Eli Sander, Vincent Roulet, Tianlin Liu, Mathieu BlondelICML 2025
