Efficient Rectified Flow for Image Fusion
Zirui Wang, Jiayi Zhang, Tianwei Guan, Yuhan Zhou, Xingyuan Li, Minjing Dong, Jinyuan Liu
Abstract
Image fusion is a fundamental and important task in computer vision, aiming to combine complementary information from different modalities to fuse images. In recent years, diffusion models have made significant developments in the field of image fusion. However, diffusion models often require complex computations and redundant inference time, which reduces the applicability of these methods. To address this issue, we propose RFfusion, an efficient one-step diffusion model for image fusion based on Rectified Flow. We incorporate Rectified Flow into the image fusion task to straighten the sampling path in the diffusion model, achieving one-step sampling without the need for additional training, while still maintaining high-quality fusion results. Furthermore, we propose a task-specific Variational Autoencoder (VAE) architecture tailored for image fusion, where the fusion operation is embedded within the latent space to further reduce computational complexity. To address the inherent discrepancy between conventional reconstruction-oriented VAE objectives and the requirements of image fusion, we introduce a two-stage training strategy. This approach facilitates the effective learning and integration of complementary information from multi-modal source images, thereby enabling the model to retain fine-grained structural details while significantly enhancing inference efficiency. Extensive experiments demonstrate that our method outperforms other state-of-the-art methods in terms of both inference speed and fusion quality. Code is available at https://github.com/zirui0625/RFfusion.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 5f3a720f-dd4b-4954-a5ed-9ac72930c55aCited by top-tier papers12
- UniFusion: A Unified Image Fusion Framework with Robust Representation and Source-Aware PreservationXingyuan Li, Songcheng Du, Yang Zou, Haoyuan Xu et al.CVPR 2026 · 6 citations
- Text-Guided Channel Perturbation and Pre-Trained Knowledge Integration for Unified Multi-Modality Image FusionXilai Li, Xiaosong Li, Weijun JiangAAAI 2026 · 4 citations
- Toward Real-world Infrared Image Super-Resolution: A Unified Autoregressive Framework and Benchmark DatasetYang Zou, Jun Ma, Zhidong Jiao, Xingyuan Li et al.CVPR 2026 · 4 citations
- Bridging Human Evaluation to Infrared and Visible Image FusionJinyuan Liu, Xingyuan Li, Qingyun Mei, HaoYuan Xu et al.CVPR 2026 · 4 citations
- Degradation-Robust Fusion: An Efficient Degradation-Aware Diffusion Framework for Multimodal Image Fusion in Arbitrary Degradation ScenariosYu Shi, Yu Liu, Zhong-Cheng Wu, Juan Cheng et al.CVPR 2026 · 4 citations
Builds on24
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Score-Based Generative Modeling through Stochastic Differential EquationsYang Song, Jascha Sohl-Dickstein, Diederik P. Kingma, Abhishek Kumar et al.ICLR 2021 · 1,270 citations
- Target-aware Dual Adversarial Learning and a Multi-scenario Multi-Modality Benchmark to Fuse Infrared and Visible for Object DetectionJinyuan Liu, Xin Fan, Zhanbo Huang, Guanyao Wu et al.CVPR 2022 · 929 citations
- FusionDN: A Unified Densely Connected Network for Image FusionHan Xu, Jiayi Ma, Zhuliang Le, Junjun Jiang et al.AAAI 2020 · 559 citations
Related papers
- Steering One-Step Diffusion Model with Fidelity-Rich Decoder for Fast Image CompressionZheng Chen, Mingde Zhou, Jinpei Guo, Jiale Yuan et al.AAAI 2026 · 1 citation
- Streaming Diffusion Model for Fast Infrared and Visible Video FusionJinyuan Liu, Ludan Sun, Tengyu Ma, Chunyan Yang et al.CVPR 2026 · 2 citations
- DRMF: Degradation-Robust Multi-Modal Image Fusion via Composable Diffusion PriorLinfeng Tang, Yuxin Deng, Xunpeng Yi, Qinglong Yan et al.ACM MM 2024 · 53 citations
- Guided and Variance-Corrected Fusion with One-shot Style Alignment for Large-Content Image GenerationShoukun Sun, Min Xian, Tiankai Yao, Fei Xu et al.AAAI 2025 · 2 citations
- Rectifying Latent Space for Generative Single-Image Reflection RemovalMingjia Li, Jin Hu, Hainuo Wang, Qiming Hu et al.CVPR 2026 · 2 citations
