Universal Approximation Using Well-Conditioned Normalizing Flows
Holden Lee, Chirag Pabbaraju, Anish Prasad Sevekari, Andrej Risteski
摘要
Normalizing flows are a widely used class of latent-variable generative models with a tractable likelihood. Affine-coupling models [Dinh et al., 2014[Dinh et al., , 2016] ] are a particularly common type of normalizing flows, for which the Jacobian of the latent-to-observable-variable transformation is triangular, allowing the likelihood to be computed in linear time. Despite the widespread usage of affine couplings, the special structure of the architecture makes understanding their representational power challenging. The question of universal approximation was only recently resolved by three parallel papers [Huang et al., 2020, Zhang et al., 2020, Koehler et al., 2020] -who showed reasonably regular distributions can be approximated arbitrarily well using affine couplings-albeit with networks with a nearly-singular Jacobian. As ill-conditioned Jacobians are an obstacle for likelihood-based training, the fundamental question remains: which distributions can be approximated using well-conditioned affine coupling flows? In this paper, we show that any log-concave distribution can be approximated using well-conditioned affine-coupling flows. In terms of proof techniques, we uncover and leverage deep connections between affine coupling architectures, underdamped Langevin dynamics (a stochastic differential equation often used to sample from Gibbs measures) and Hénon maps (a structured dynamical system that appears in the study of symplectic diffeomorphisms). Our results also inform the practice of training affine couplings: we approximate a padded version of the input distribution with iid Gaussians-a strategy which Koehler et al. [2020] empirically observed to result in better-conditioned flows, but had hitherto no theoretical grounding. Our proof can thus be seen as providing theoretical evidence for the benefits of Gaussian padding when training normalizing flows.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- On the Universality of Volume-Preserving and Coupling-Based Normalizing FlowsFelix Draxler, Stefan Wahl, Christoph Schnörr, Ullrich KötheICML 2024 · 被引用 19 次
- Scalable Equilibrium Sampling with Sequential Boltzmann GeneratorsCharlie B. Tan, Joey Bose, Chen Lin, Leon Klein 等ICML 2025
它引用的顶会 Paper6
- Score-Based Generative Modeling through Stochastic Differential EquationsYang Song, Jascha Sohl-Dickstein, Diederik P. Kingma, Abhishek Kumar 等ICLR 2021 · 被引用 1,270 次
- Relaxing Bijectivity Constraints with Continuously Indexed Normalising FlowsRobert Cornish, Anthony L. Caterini, George Deligiannidis, Arnaud DoucetICML 2020 · 被引用 141 次
- Coupling-based Invertible Neural Networks Are Universal Diffeomorphism ApproximatorsTakeshi Teshima, Isao Ishikawa, Koichi Tojo, Kenta Oono 等NeurIPS 2020 · 被引用 129 次
- Approximation Capabilities of Neural ODEs and Invertible Residual NetworksHan Zhang, Xi Gao, Jacob Unterman, Tom ArodzICML 2020 · 被引用 114 次
- VFlow: More Expressive Generative Flows with Variational Data AugmentationJianfei Chen, Cheng Lu, Biqi Chenli, Jun Zhu 等ICML 2020 · 被引用 64 次
相关 Paper
- Representational aspects of depth and conditioning in normalizing flowsFrederic Koehler, Viraj Mehta, Andrej RisteskiICML 2021 · 被引用 29 次
- Whitening Convergence Rate of Coupling-based Normalizing FlowsFelix Draxler, Christoph Schnörr, Ullrich KötheNeurIPS 2022 · 被引用 7 次
- Self Normalizing FlowsT. Anderson Keller, Jorn W. T. Peters, Priyank Jaini, Emiel Hoogeboom 等ICML 2021 · 被引用 14 次
- On the Robustness of Normalizing Flows for Inverse Problems in ImagingSeongmin Hong, Inbum Park, Se Young ChunICCV 2023 · 被引用 9 次
- Latent Variable Modelling with Hyperbolic Normalizing FlowsAvishek Joey Bose, Ariella Smofsky, Renjie Liao, Prakash Panangaden 等ICML 2020 · 被引用 76 次
