Representational aspects of depth and conditioning in normalizing flows
Frederic Koehler, Viraj Mehta, Andrej Risteski
摘要
Normalizing flows are among the most popular paradigms in generative modeling, especially for images, primarily because we can efficiently evaluate the likelihood of a data point. Normalizing flows also come with difficulties: models which produce good samples typically need to be extremely deep -- which comes with accompanying vanishing/exploding gradient problems. Relatedly, they are often poorly conditioned since typical training data like images intuitively are lower-dimensional, and the learned maps often have Jacobians that are close to being singular. In our paper, we tackle representational aspects around depth and conditioning of normalizing flows -- both for general invertible architectures, and for a particular common architecture -- affine couplings. For general invertible architectures, we prove that invertibility comes at a cost in terms of depth: we show examples where a much deeper normalizing flow model may need to be used to match the performance of a non-invertible generator. For affine couplings, we first show that the choice of partitions isn't a likely bottleneck for depth: we show that any invertible linear map (and hence a permutation) can be simulated by a constant number of affine coupling layers, using a fixed partition. This shows that the extra flexibility conferred by 1x1 convolution layers, as in GLOW, can in principle be simulated by increasing the size by a constant factor. Next, in terms of conditioning, we show that affine couplings are universal approximators -- provided the Jacobian of the model is allowed to be close to singular. We furthermore empirically explore the benefit of different kinds of padding -- a common strategy for improving conditioning -- on both synthetic and real-life datasets.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper11
- On the Universality of Volume-Preserving and Coupling-Based Normalizing FlowsFelix Draxler, Stefan Wahl, Christoph Schnörr, Ullrich KötheICML 2024 · 被引用 19 次
- Low Complexity Homeomorphic Projection to Ensure Neural-Network Solution Feasibility for Optimization over (Non-)Convex SetEnming Liang, Minghua Chen, Steven H. LowICML 2023 · 被引用 18 次
- Universal Approximation Using Well-Conditioned Normalizing FlowsHolden Lee, Chirag Pabbaraju, Anish Prasad Sevekari, Andrej RisteskiNeurIPS 2021 · 被引用 15 次
- MGF: Mixed Gaussian Flow for Diverse Trajectory PredictionJiahe Chen, Jinkun Cao, Dahua Lin, Kris Kitani 等NeurIPS 2024 · 被引用 11 次
- Whitening Convergence Rate of Coupling-based Normalizing FlowsFelix Draxler, Christoph Schnörr, Ullrich KötheNeurIPS 2022 · 被引用 7 次
它引用的顶会 Paper2
- Coupling-based Invertible Neural Networks Are Universal Diffeomorphism ApproximatorsTakeshi Teshima, Isao Ishikawa, Koichi Tojo, Kenta Oono 等NeurIPS 2020 · 被引用 129 次
- Approximation Capabilities of Neural ODEs and Invertible Residual NetworksHan Zhang, Xi Gao, Jacob Unterman, Tom ArodzICML 2020 · 被引用 114 次
相关 Paper
- Densely connected normalizing flowsMatej Grcic, Ivan Grubisic, Sinisa SegvicNeurIPS 2021 · 被引用 67 次
- Self Normalizing FlowsT. Anderson Keller, Jorn W. T. Peters, Priyank Jaini, Emiel Hoogeboom 等ICML 2021 · 被引用 14 次
- On the Robustness of Normalizing Flows for Inverse Problems in ImagingSeongmin Hong, Inbum Park, Se Young ChunICCV 2023 · 被引用 9 次
- Generative Flows with Matrix ExponentialChangyi Xiao, Ligang LiuICML 2020 · 被引用 10 次
- Flowification: Everything is a normalizing flowBálint Máté, Samuel Klein, Tobias Golling, François FleuretNeurIPS 2022 · 被引用 9 次
