On the difficulty of unbiased alpha divergence minimization
Tomas Geffner, Justin Domke
摘要
Several approximate inference algorithms have been proposed to minimize an alpha-divergence between an approximating distribution and a target distribution. Many of these algorithms introduce bias, the magnitude of which becomes problematic in high dimensions. Other algorithms are unbiased. These often seem to suffer from high variance, but little is rigorously known. In this work we study unbiased methods for alpha-divergence minimization through the Signal-to-Noise Ratio (SNR) of the gradient estimator. We study several representative scenarios where strong analytical results are possible, such as fully-factorized or Gaussian distributions. We find that when alpha is not zero, the SNR worsens exponentially in the dimensionality of the problem. This casts doubt on the practicality of these methods. We empirically confirm these theoretical results.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper11
- Challenges and Opportunities in High Dimensional Variational InferenceAkash Kumar Dhaka, Alejandro Catalina, Manushi Welandawe, Michael Riis Andersen 等NeurIPS 2021 · 被引用 54 次
- Path-Gradient Estimators for Continuous Normalizing FlowsLorenz Vaitl, Kim Andrea Nicoli, Shinichi Nakajima, Pan KesselICML 2022 · 被引用 14 次
- Markov Chain Score Ascent: A Unifying Framework of Variational Inference with Markovian GradientsKyurae Kim, Jisu Oh, Jacob R. Gardner, Adji Bousso Dieng 等NeurIPS 2022 · 被引用 12 次
- Mixture weights optimisation for Alpha-Divergence Variational InferenceKamélia Daudel, Randal DoucNeurIPS 2021 · 被引用 11 次
- Large Language BayesJustin DomkeNeurIPS 2025 · 被引用 10 次
相关 Paper
- Asymptotics of Alpha-Divergence Variational Inference Algorithms with Exponential FamiliesFrançois Bertholom, Randal Douc, François RoueffNeurIPS 2024 · 被引用 1 次
- Dimension-free Private Mean Estimation for Anisotropic DistributionsYuval Dagan, Michael I. Jordan, Xuelin Yang, Lydia Zakynthinou 等NeurIPS 2024 · 被引用 7 次
- Entangled Mean Estimation in High DimensionsIlias Diakonikolas, Daniel M. Kane, Sihan Liu, Thanasis PittasSTOC 2025 · 被引用 1 次
- Computing Divergences between Discrete Decomposable ModelsLoong Kuan Lee, Nico Piatkowski, François Petitjean, Geoffrey I. WebbAAAI 2023 · 被引用 4 次
- Understanding the Variance Collapse of SVGD in High DimensionsJimmy Ba, Murat A. Erdogdu, Marzyeh Ghassemi, Shengyang Sun 等ICLR 2022 · 被引用 35 次
