Densely connected normalizing flows
Matej Grcic, Ivan Grubisic, Sinisa Segvic
Abstract
Normalizing flows are bijective mappings between inputs and latent representations with a fully factorized distribution. They are very attractive due to exact likelihood evaluation and efficient sampling. However, their effective capacity is often insufficient since the bijectivity constraint limits the model width. We address this issue by incrementally padding intermediate representations with noise. We precondition the noise in accordance with previous invertible units, which we describe as crossunit coupling. Our invertible glow-like modules increase the model expressivity by fusing a densely connected block with Nyström self-attention. We refer to our architecture as DenseFlow since both cross-unit and intra-module couplings rely on dense connectivity. Experiments show significant improvements due to the proposed contributions and reveal state-of-the-art density estimation under moderate computing budgets. 1
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 6ddb284f-2730-4656-b59e-deeb3bde5564Cited by top-tier papers13
- Consistency ModelsYang Song, Prafulla Dhariwal, Mark Chen, Ilya SutskeverICML 2023 · 1,720 citations
- Soft Truncation: A Universal Training Technique of Score-based Diffusion Model for High Precision Score EstimationDongjun Kim, Seungjae Shin, Kyungwoo Song, Wanmo Kang et al.ICML 2022 · 115 citations
- Stable, Fast and Accurate: Kernelized Attention with Relative Positional EncodingShengjie Luo, Shanda Li, Tianle Cai, Di He et al.NeurIPS 2021 · 66 citations
- Maximum Likelihood Training of Implicit Nonlinear Diffusion ModelDongjun Kim, Byeonghu Na, Se Jung Kwon, Dongsoo Lee et al.NeurIPS 2022 · 61 citations
- Transformer-VQ: Linear-Time Transformers via Vector QuantizationLucas D. LingleICLR 2024 · 30 citations
Builds on14
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 11,743 citations
- Improved Denoising Diffusion Probabilistic ModelsAlexander Quinn Nichol, Prafulla DhariwalICML 2021 · 5,234 citations
- NVAE: A Deep Hierarchical Variational AutoencoderArash Vahdat, Jan KautzNeurIPS 2020 · 1,141 citations
Related papers
- Invertible DenseNets with Concatenated LipSwishYura Perugachi-Diaz, Jakub M. Tomczak, Sandjai BhulaiNeurIPS 2021 · 29 citations
- Representational aspects of depth and conditioning in normalizing flowsFrederic Koehler, Viraj Mehta, Andrej RisteskiICML 2021 · 29 citations
- Gradient Boosted Normalizing FlowsRobert A. Giaquinto, Arindam BanerjeeNeurIPS 2020 · 11 citations
- NanoFlow: Scalable Normalizing Flows with Sublinear Parameter ComplexitySang-gil Lee, Sungwon Kim, Sungroh YoonNeurIPS 2020 · 20 citations
- Self Normalizing FlowsT. Anderson Keller, Jorn W. T. Peters, Priyank Jaini, Emiel Hoogeboom et al.ICML 2021 · 14 citations
