Generative Flows with Invertible Attentions
Rhea Sanjay Sukthanker, Zhiwu Huang, Suryansh Kumar, Radu Timofte, Luc Van Gool
摘要
Flow-based generative models have shown an excellent ability to explicitly learn the probability density function of data via a sequence of invertible transformations. Yet, learning attentions in generative flows remains understudied, while it has made breakthroughs in other domains. To fill the gap, this paper introduces two types of invertible attention mechanisms, i.e., map-based and transformer-based attentions, for both unconditional and conditional generative flows. The key idea is to exploit a masked scheme of these two attentions to learn long-range data dependencies in the context of generative flows. The masked scheme allows for invertible attention modules with tractable Jacobian determinants, enabling its seamless integration at any positions of the flow-based models. The proposed attention mechanisms lead to more efficient generative flows, due to their capability of modeling the long-term data dependencies. Evaluation on multiple image synthesis tasks shows that the proposed attention flows result in efficient models and compare favorably against the state-of-the-art unconditional and conditional generative flows.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- On the Robustness of Normalizing Flows for Inverse Problems in ImagingSeongmin Hong, Inbum Park, Se Young ChunICCV 2023 · 被引用 9 次
- MANGO: Multimodal Attention-based Normalizing Flow Approach to Fusion LearningThanh-Dat Truong, Christophe Bobda, Nitin Agarwal, Khoa LuuNeurIPS 2025 · 被引用 6 次
- Probabilistic Forecasting of Irregularly Sampled Time Series with Missing Values via Conditional Normalizing FlowsVijaya Krishna Yalavarthi, Randolf Scholz, Stefan Born, Lars Schmidt-ThiemeAAAI 2025 · 被引用 6 次
- Neural Diffeomorphic Non-uniform B-spline FlowsSeongmin Hong, Se Young ChunAAAI 2023 · 被引用 3 次
- From Softmax to Score: Transformers Can Effectively Implement In-Context Denoising StepsPaul Rosu, Lawrence Carin, Xiang ChengNeurIPS 2025 · 被引用 3 次
它引用的顶会 Paper12
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu 等ICCV 2021 · 被引用 31,683 次
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- Attention Augmented Convolutional NetworksIrwan Bello, Barret Zoph, Quoc Le, Ashish Vaswani 等ICCV 2019 · 被引用 1,149 次
- Attention is not all you need: pure attention loses rank doubly exponentially with depthYihe Dong, Jean-Baptiste Cordonnier, Andreas LoukasICML 2021 · 被引用 522 次
- TransGAN: Two Pure Transformers Can Make One Strong GAN, and That Can Scale UpYifan Jiang, Shiyu Chang, Zhangyang WangNeurIPS 2021 · 被引用 515 次
相关 Paper
- Normalizing Flows With Multi-Scale Autoregressive PriorsApratim Bhattacharyya, Shweta Mahajan, Mario Fritz, Bernt Schiele 等CVPR 2020
- Attentive Normalization for Conditional Image GenerationYi Wang, Ying-Cong Chen, Xiangyu Zhang, Jian Sun 等CVPR 2020
- Evolving Attention with Residual ConvolutionsYujing Wang, Yaming Yang, Jiangang Bai, Mingliang Zhang 等ICML 2021 · 被引用 43 次
- Video Frame Interpolation with Flow TransformerPan Gao, Haoyue Tian, Jie QinACM MM 2023 · 被引用 4 次
- MaskSketch: Unpaired Structure-guided Masked Image GenerationDina Bashkirova, José Lezama, Kihyuk Sohn, Kate Saenko 等CVPR 2023
