Efficient Learning of Generative Models via Finite-Difference Score Matching
Tianyu Pang, Taufik Xu, Chongxuan Li, Yang Song, Stefano Ermon, Jun Zhu
摘要
Several machine learning applications involve the optimization of higher-order derivatives (e.g., gradients of gradients) during training, which can be expensive with respect to memory and computation even with automatic differentiation. As a typical example in generative modeling, score matching (SM) involves the optimization of the trace of a Hessian. To improve computing efficiency, we rewrite the SM objective and its variants in terms of directional derivatives, and present a generic strategy to efficiently approximate any-order directional derivative with finite difference (FD). Our approximation only involves function evaluations, which can be executed in parallel, and no gradient computations. Thus, it reduces the total computational cost while also improving numerical stability. We provide two instantiations by reformulating variants of SM objectives into the FD forms. Empirically, we demonstrate that our methods produce results comparable to the gradient-based counterparts while being much more computationally efficient.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper31
- Score-Based Generative Modeling through Stochastic Differential EquationsYang Song, Jascha Sohl-Dickstein, Diederik P. Kingma, Abhishek Kumar 等ICLR 2021 · 被引用 1,270 次
- Score-based Generative Modeling in Latent SpaceArash Vahdat, Karsten Kreis, Jan KautzNeurIPS 2021 · 被引用 903 次
- Robustness and Accuracy Could Be Reconcilable by (Proper) DefinitionTianyu Pang, Min Lin, Xiao Yang, Jun Zhu 等ICML 2022 · 被引用 168 次
- Concrete Score Matching: Generalized Score Matching for Discrete DataChenlin Meng, Kristy Choi, Jiaming Song, Stefano ErmonNeurIPS 2022 · 被引用 168 次
- Estimating High Order Gradients of the Data Distribution by DenoisingChenlin Meng, Yang Song, Wenzhe Li, Stefano ErmonNeurIPS 2021 · 被引用 85 次
它引用的顶会 Paper5
- Your classifier is secretly an energy based model and you should treat it like oneWill Grathwohl, Kuan-Chieh Wang, Jörn-Henrik Jacobsen, David Duvenaud 等ICLR 2020 · 被引用 643 次
- On the Anatomy of MCMC-Based Maximum Likelihood Learning of Energy-Based ModelsErik Nijkamp, Mitch Hill, Tian Han, Song-Chun Zhu 等AAAI 2020 · 被引用 182 次
- Nonparametric Score EstimatorsYuhao Zhou, Jiaxin Shi, Jun ZhuICML 2020 · 被引用 30 次
- SUMO: Unbiased Estimation of Log Marginal Probability for Latent Variable ModelsYucen Luo, Alex Beatson, Mohammad Norouzi, Jun Zhu 等ICLR 2020 · 被引用 29 次
- Bi-level Score Matching for Learning Energy-based Latent Variable ModelsFan Bao, Chongxuan Li, Taufik Xu, Hang Su 等NeurIPS 2020 · 被引用 16 次
相关 Paper
- Efficiently Access Diffusion Fisher: Within the Outer Product Span SpaceFangyikang Wang, Hubery Yin, Shaobin Zhuang, Huminhao Zhu 等ICML 2025
- Efficient Score Matching with Deep Equilibrium LayersYuhao Huang, Qingsong Wang, Akwum Onwunta, Bao WangICLR 2024 · 被引用 4 次
- Variational (Gradient) Estimate of the Score Function in Energy-based Latent Variable ModelsFan Bao, Kun Xu, Chongxuan Li, Lanqing Hong 等ICML 2021 · 被引用 10 次
- MissScore: High-Order Score Estimation in the Presence of Missing DataWenqin Liu, Haoze Hou, Erdun Gao, Biwei Huang 等ICML 2025
- Maximum Likelihood Training for Score-based Diffusion ODEs by High Order Denoising Score MatchingCheng Lu, Kaiwen Zheng, Fan Bao, Jianfei Chen 等ICML 2022 · 被引用 109 次
