Efficient Learning of Generative Models via Finite-Difference Score Matching
Tianyu Pang, Taufik Xu, Chongxuan Li, Yang Song, Stefano Ermon, Jun Zhu
Abstract
Several machine learning applications involve the optimization of higher-order derivatives (e.g., gradients of gradients) during training, which can be expensive with respect to memory and computation even with automatic differentiation. As a typical example in generative modeling, score matching (SM) involves the optimization of the trace of a Hessian. To improve computing efficiency, we rewrite the SM objective and its variants in terms of directional derivatives, and present a generic strategy to efficiently approximate any-order directional derivative with finite difference (FD). Our approximation only involves function evaluations, which can be executed in parallel, and no gradient computations. Thus, it reduces the total computational cost while also improving numerical stability. We provide two instantiations by reformulating variants of SM objectives into the FD forms. Empirically, we demonstrate that our methods produce results comparable to the gradient-based counterparts while being much more computationally efficient.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext e5eb173a-11ba-4b26-8abe-069134ed8620Cited by top-tier papers31
- Score-Based Generative Modeling through Stochastic Differential EquationsYang Song, Jascha Sohl-Dickstein, Diederik P. Kingma, Abhishek Kumar et al.ICLR 2021 · 1,270 citations
- Score-based Generative Modeling in Latent SpaceArash Vahdat, Karsten Kreis, Jan KautzNeurIPS 2021 · 903 citations
- Robustness and Accuracy Could Be Reconcilable by (Proper) DefinitionTianyu Pang, Min Lin, Xiao Yang, Jun Zhu et al.ICML 2022 · 168 citations
- Concrete Score Matching: Generalized Score Matching for Discrete DataChenlin Meng, Kristy Choi, Jiaming Song, Stefano ErmonNeurIPS 2022 · 168 citations
- Estimating High Order Gradients of the Data Distribution by DenoisingChenlin Meng, Yang Song, Wenzhe Li, Stefano ErmonNeurIPS 2021 · 85 citations
Builds on5
- Your classifier is secretly an energy based model and you should treat it like oneWill Grathwohl, Kuan-Chieh Wang, Jörn-Henrik Jacobsen, David Duvenaud et al.ICLR 2020 · 643 citations
- On the Anatomy of MCMC-Based Maximum Likelihood Learning of Energy-Based ModelsErik Nijkamp, Mitch Hill, Tian Han, Song-Chun Zhu et al.AAAI 2020 · 182 citations
- Nonparametric Score EstimatorsYuhao Zhou, Jiaxin Shi, Jun ZhuICML 2020 · 30 citations
- SUMO: Unbiased Estimation of Log Marginal Probability for Latent Variable ModelsYucen Luo, Alex Beatson, Mohammad Norouzi, Jun Zhu et al.ICLR 2020 · 29 citations
- Bi-level Score Matching for Learning Energy-based Latent Variable ModelsFan Bao, Chongxuan Li, Taufik Xu, Hang Su et al.NeurIPS 2020 · 16 citations
Related papers
- Efficiently Access Diffusion Fisher: Within the Outer Product Span SpaceFangyikang Wang, Hubery Yin, Shaobin Zhuang, Huminhao Zhu et al.ICML 2025
- Efficient Score Matching with Deep Equilibrium LayersYuhao Huang, Qingsong Wang, Akwum Onwunta, Bao WangICLR 2024 · 4 citations
- Variational (Gradient) Estimate of the Score Function in Energy-based Latent Variable ModelsFan Bao, Kun Xu, Chongxuan Li, Lanqing Hong et al.ICML 2021 · 10 citations
- MissScore: High-Order Score Estimation in the Presence of Missing DataWenqin Liu, Haoze Hou, Erdun Gao, Biwei Huang et al.ICML 2025
- Maximum Likelihood Training for Score-based Diffusion ODEs by High Order Denoising Score MatchingCheng Lu, Kaiwen Zheng, Fan Bao, Jianfei Chen et al.ICML 2022 · 109 citations
