Generative Flows with Invertible Attentions
Rhea Sanjay Sukthanker, Zhiwu Huang, Suryansh Kumar, Radu Timofte, Luc Van Gool
Abstract
Flow-based generative models have shown an excellent ability to explicitly learn the probability density function of data via a sequence of invertible transformations. Yet, learning attentions in generative flows remains understudied, while it has made breakthroughs in other domains. To fill the gap, this paper introduces two types of invertible attention mechanisms, i.e., map-based and transformer-based attentions, for both unconditional and conditional generative flows. The key idea is to exploit a masked scheme of these two attentions to learn long-range data dependencies in the context of generative flows. The masked scheme allows for invertible attention modules with tractable Jacobian determinants, enabling its seamless integration at any positions of the flow-based models. The proposed attention mechanisms lead to more efficient generative flows, due to their capability of modeling the long-term data dependencies. Evaluation on multiple image synthesis tasks shows that the proposed attention flows result in efficient models and compare favorably against the state-of-the-art unconditional and conditional generative flows.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 0e5b1983-31b7-4fe5-bdea-ec852ccab84eCited by top-tier papers6
- On the Robustness of Normalizing Flows for Inverse Problems in ImagingSeongmin Hong, Inbum Park, Se Young ChunICCV 2023 · 9 citations
- MANGO: Multimodal Attention-based Normalizing Flow Approach to Fusion LearningThanh-Dat Truong, Christophe Bobda, Nitin Agarwal, Khoa LuuNeurIPS 2025 · 6 citations
- Probabilistic Forecasting of Irregularly Sampled Time Series with Missing Values via Conditional Normalizing FlowsVijaya Krishna Yalavarthi, Randolf Scholz, Stefan Born, Lars Schmidt-ThiemeAAAI 2025 · 6 citations
- Neural Diffeomorphic Non-uniform B-spline FlowsSeongmin Hong, Se Young ChunAAAI 2023 · 3 citations
- From Softmax to Score: Transformers Can Effectively Implement In-Context Denoising StepsPaul Rosu, Lawrence Carin, Xiang ChengNeurIPS 2025 · 3 citations
Builds on12
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu et al.ICCV 2021 · 31,683 citations
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Attention Augmented Convolutional NetworksIrwan Bello, Barret Zoph, Quoc Le, Ashish Vaswani et al.ICCV 2019 · 1,149 citations
- Attention is not all you need: pure attention loses rank doubly exponentially with depthYihe Dong, Jean-Baptiste Cordonnier, Andreas LoukasICML 2021 · 522 citations
- TransGAN: Two Pure Transformers Can Make One Strong GAN, and That Can Scale UpYifan Jiang, Shiyu Chang, Zhangyang WangNeurIPS 2021 · 515 citations
Related papers
- Normalizing Flows With Multi-Scale Autoregressive PriorsApratim Bhattacharyya, Shweta Mahajan, Mario Fritz, Bernt Schiele et al.CVPR 2020
- Attentive Normalization for Conditional Image GenerationYi Wang, Ying-Cong Chen, Xiangyu Zhang, Jian Sun et al.CVPR 2020
- Evolving Attention with Residual ConvolutionsYujing Wang, Yaming Yang, Jiangang Bai, Mingliang Zhang et al.ICML 2021 · 43 citations
- Video Frame Interpolation with Flow TransformerPan Gao, Haoyue Tian, Jie QinACM MM 2023 · 4 citations
- MaskSketch: Unpaired Structure-guided Masked Image GenerationDina Bashkirova, José Lezama, Kihyuk Sohn, Kate Saenko et al.CVPR 2023
